Projects in Awesome Lists tagged with reasoning-models
A curated list of projects in awesome lists tagged with reasoning-models .
https://zilliztech.github.io/deep-searcher/
Open Source Deep Research Alternative to Reason and Search on Private Data. Written in Python.
agent agentic-rag claude deep-research deepseek deepseek-r1 grok grok3 llama4 llm milvus openai qwen3 rag reasoning-models vector-database zilliz
Last synced: 22 Jul 2025
https://github.com/iaar-shanghai/xverify
xVerify: Efficient Answer Verifier for Reasoning Model Evaluations
benchmark cc-by-nc-nd-4 chatgpt deepseek-math evaluation judge-model llm llm-as-a-judge math-verify open-compass open-r1 reasoning-models regex reliability reliability-tools xverify
Last synced: 07 Oct 2025
https://github.com/MiniMax-AI/MiniMax-M1
MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model.
large-language-models llm minimax-m1 reasoning-models
Last synced: 22 Jun 2025
https://github.com/ukplab/acl2025-diverse-cot
Code for the 2025 ACL publication "Fine-Tuning on Diverse Reasoning Chains Drives Within-Inference CoT Refinement in LLMs"
chain-of-thought cot large-reasoning-models lrm reasoning-models
Last synced: 22 Jul 2025
https://github.com/sinanuozdemir/oreilly-agi
Explore the evolution of AGI through historical context, reasoning models, and agent systems, while gaining hands-on experience with cutting-edge models like Claude 4, DeepSeek-R1, and OpenAI's o3. Learn to critically evaluate AGI benchmarks, understand their limitations, and identify where current models excel or struggle in reasoning tasks.
agents agi ai-agents artifical-general-inteligence reasoning-models
Last synced: 07 Mar 2026
https://github.com/mrorigo/agentic-deep-graph-reasoning
Agentic Deep Graph Reasoning Implementation
ai-learning entity-extraction knowledge-distillation knowledge-graph reasoning-models
Last synced: 20 Mar 2025
https://github.com/sshh12/state-sandbox
State Sandbox is an experimental game for socioeconomic simulation. It uses Large Language Models (o3-mini) to simulate the world and complex policy impacts.
ai-games civilization nation-states o1 o3-mini reasoning-models socioeconomics
Last synced: 17 Mar 2026
https://github.com/codelion/pts
Pivotal Token Search
dataset-generation direct-preference-optimization dpo llm llm-inference llm-steering mech-interp phi-4 phi-4-mini phi4 phi4-mini pivotal-token-search pivotal-tokens reasoning-agent reasoning-language-models reasoning-models sae sparse-autoencoder steering-vector tokens
Last synced: 10 Jun 2025
https://github.com/contactvaibhavi/hyper-personalised-agent
Hyper-personalised agentic system to aggregate, reason and plan over multiple input streams
agent planning-algorithms reasoning-models retrieval-augmented-generation
Last synced: 11 Jun 2026
https://github.com/akhilpandey95/s1
Experiments on test-time scaling approaches for reasoning LM's to enforce better <think> or <wait> capabilities.
deepseek-r1 inference reasoning-models test-time-computation ttc
Last synced: 03 May 2026
https://github.com/shaheennabi/rlvr_grpo-experiment-with-math500
A small experiment repository comparing a base reasoning model against RLVR-GRPO checkpoints on the Math500 dataset. It includes evaluation results, short-form observations, and a local temp_clone of the full open-posttraining-system codebase for reference.
evaluating-models grpo-checkpoint math500 open-posttraining-system policy-optimization post-training reasoning-models reinforcement-learning rlvr-grpo sparse-rewards
Last synced: 18 Jun 2026
https://github.com/kaicheng001/awesome-r1
A curated list of research papers, models, and resources related to R1-style reasoning models following DeepSeek-R1's breakthrough in January 2025.
awesome deepseek-r1 llm lmm mllm r1 reasoning-models reward-model thinking vlm
Last synced: 22 Jun 2025
https://github.com/angelnicolasc/meridian
Phase-aware vLLM scheduler for reasoning models: output-first dispatch, entropy-gated think termination, tiered KV eviction, and TTOT-focused benchmarking.
cuda inference kv-cache llm observability pyo3 python reasoning-models rust scheduler vllm
Last synced: 27 Jun 2026