Skip to content
TokenPulse AI
Advertisement
Header banner · 728 × 90Ad network placeholder - ready for your script
Trending AI Intelligence

Trending GitHub AI repos
and record-breaking papers

A curated, source-linked snapshot of the open-source AI repositories and research reports developers are watching - from DeepSeek and Qwen to frontier agents, reasoning, and evaluation benchmarks.

Trending GitHub AI Repositories

Trending GitHub AI repositories

A curated snapshot of notable, high-activity open-source AI repositories. Star counts and momentum are reference values that change constantly - open each repo to verify current numbers.

Filter:
huggingface/transformersllms
Steady

State-of-the-art machine learning library for PyTorch, TensorFlow, and JAX.

~141.0kchecked Aug 2026
View
ollama/ollamainference
High

Run large language models locally with a simple CLI and REST API.

~124.0kchecked Aug 2026
View
langchain-ai/langchaintooling
High

Framework for building LLM applications with composable chains, tools, and agents.

~98.0kchecked Aug 2026
View
ggml-org/llama.cppinference
Steady

LLM inference in plain C/C++ with minimal dependencies and broad hardware support.

~72.0kchecked Aug 2026
View
comfyanonymous/ComfyUIvision
High

Powerful and modular diffusion model GUI and backend for image generation.

~63.0kchecked Aug 2026
View
open-webui/open-webuitooling
High

User-friendly WebUI for LLMs with a rich feature set and offline support.

~62.0kchecked Aug 2026
View
langgenius/difytooling
High

Open-source LLM app development platform for building AI workflows and agents.

~61.0kchecked Aug 2026
View
openai/openai-cookbookllms
Steady

Examples and guides for using the OpenAI API, from prompts to function calling.

~61.0kchecked Aug 2026
View
facebookresearch/segment-anythingvision
Steady

Segment Anything Model (SAM): promptable segmentation for images and video.

~49.0kchecked Aug 2026
View
microsoft/autogenagents
Rising

Multi-agent conversation framework for building next-gen AI applications.

~42.0kchecked Aug 2026
View
vllm-project/vllminference
High

High-throughput, memory-efficient inference and serving engine for LLMs.

~41.0kchecked Aug 2026
View
run-llama/llama_indexdata
Steady

Data framework for connecting LLMs to your own documents and data sources.

~39.0kchecked Aug 2026
View
Aider-AI/aidertooling
Rising

AI pair programming in your terminal, editing code across your repo.

~26.0kchecked Aug 2026
View
BerriAI/litellminference
High

Python SDK, proxy server, and gateway for calling 100+ LLMs with one interface.

~16.0kchecked Aug 2026
View
langchain-ai/langgraphagents
High

Low-level orchestration framework for building stateful, multi-agent LLM apps.

~11.0kchecked Aug 2026
View
Sponsor
AI Research Papers & Reports

Top AI research papers & benchmarks

Record-breaking and widely cited research across reasoning, frontier agents, inference efficiency, multimodal systems, and evaluation. Each item links to its official source.

Filter:
inferenceOfficial source· Dec 2024

DeepSeek-V3 Technical Report

Why it matters: An open-weights model with a Mixture-of-Experts architecture and a cost-efficient training pipeline that reset expectations for frontier open models.

Publisher: DeepSeek-AI
Source
reasoningOfficial source· Jan 2025

DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Why it matters: Showed that reasoning can be elicited with reinforcement learning, and open-sourced the reasoning model and its distillation recipes.

Publisher: DeepSeek-AI
Source
llmsOfficial source· 2025

Qwen2.5 Technical Report

Why it matters: A family of open models spanning dense and MoE sizes with strong multilingual and coding performance.

Publisher: Qwen Team, Alibaba
Source
llmsOfficial source· 2025

Qwen3 Technical Report

Why it matters: Next-generation open model family with hybrid thinking modes and native tool use.

Publisher: Qwen Team, Alibaba
Source
reasoningOfficial source· 2024

OpenAI o1 System Card

Why it matters: Introduced a model trained with reinforcement learning to think before answering, changing the reasoning paradigm.

Publisher: OpenAI
Source
agentsOfficial source· 2025

Claude 3.7 Sonnet System Card

Why it matters: A hybrid reasoning model with extended thinking plus computer-use tooling for agentic workflows.

Publisher: Anthropic
Source
agentsOfficial source· 2024

Claude 3.5 Sonnet System Card

Why it matters: Introduced computer use, a step toward models that operate tools and interfaces.

Publisher: Anthropic
Source
multimodalOfficial source· 2024

Gemini 1.5 Technical Report

Why it matters: A long-context (up to 1M tokens) multimodal model that set a new bar for context handling.

Publisher: Google DeepMind
Source
multimodalOfficial source· Dec 2024

Gemini 2.0

Why it matters: An agentic-era multimodal model with native tool use and real-time capabilities.

Publisher: Google DeepMind
Source
llmsOfficial source· 2024

The Llama 3 Herd of Models

Why it matters: An open-weight flagship family covering text, vision, and speech, widely used as a baseline.

Publisher: Meta AI
Source
inferenceOfficial source· 2023

PagedAttention: Efficient Memory Management for LLM Serving

Why it matters: The memory-management technique behind vLLM that made high-throughput serving practical.

Publisher: vLLM team
Source
evaluationOfficial source· 2024

MMLU-Pro

Why it matters: A more challenging, reasoning-heavy successor to MMLU for evaluating knowledge and reasoning.

Publisher: TIGER-Lab
Source
evaluationOfficial source· 2023

SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Why it matters: The standard benchmark for measuring how well models can fix real software issues.

Publisher: Princeton NLP
Source
evaluationOfficial source· 2021

HumanEval

Why it matters: A widely cited benchmark for code generation that most coding models report.

Publisher: OpenAI
Source
evaluationOfficial source· 2023

LMSYS Chatbot Arena

Why it matters: A crowd-sourced, Elo-based leaderboard for comparing LLMs by human preference.

Publisher: LMSYS / LM Arena
Source
evaluationOfficial source· 2022

HELM: Holistic Evaluation of Language Models

Why it matters: A broad, multi-metric evaluation framework for transparent model comparison.

Publisher: Stanford CRFM
Source
evaluationOfficial source· 2023

GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Why it matters: A hard, expert-level benchmark used to gauge frontier reasoning.

Publisher: NYU
Source
agentsOfficial source· 2023

Gorilla: Large Language Models Connected with Massive APIs

Why it matters: A benchmark and model family for API and tool calling, foundational for agent tool use.

Publisher: UC Berkeley
Source
Sponsor
Data freshness & methodology

Star counts, growth, rankings, and trend status on this page are curated reference snapshots, not live data. They change frequently and should be verified at the source before you rely on them. We do not fetch live numbers without an API key, and we never fabricate figures. Use the "View" and "Source" links on each card to check current values directly on GitHub, arXiv, or the publisher's site.