An open API service indexing awesome lists of open source software.

Projects in Awesome Lists tagged with reasoning-models

A curated list of projects in awesome lists tagged with reasoning-models .

https://zilliztech.github.io/deep-searcher/

Open Source Deep Research Alternative to Reason and Search on Private Data. Written in Python.

agent agentic-rag claude deep-research deepseek deepseek-r1 grok grok3 llama4 llm milvus openai qwen3 rag reasoning-models vector-database zilliz

Last synced: 22 Jul 2025

https://github.com/MiniMax-AI/MiniMax-M1

MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model.

large-language-models llm minimax-m1 reasoning-models

Last synced: 22 Jun 2025

https://github.com/ukplab/acl2025-diverse-cot

Code for the 2025 ACL publication "Fine-Tuning on Diverse Reasoning Chains Drives Within-Inference CoT Refinement in LLMs"

chain-of-thought cot large-reasoning-models lrm reasoning-models

Last synced: 22 Jul 2025

https://github.com/sinanuozdemir/oreilly-agi

Explore the evolution of AGI through historical context, reasoning models, and agent systems, while gaining hands-on experience with cutting-edge models like Claude 4, DeepSeek-R1, and OpenAI's o3. Learn to critically evaluate AGI benchmarks, understand their limitations, and identify where current models excel or struggle in reasoning tasks.

agents agi ai-agents artifical-general-inteligence reasoning-models

Last synced: 07 Mar 2026

https://github.com/sshh12/state-sandbox

State Sandbox is an experimental game for socioeconomic simulation. It uses Large Language Models (o3-mini) to simulate the world and complex policy impacts.

ai-games civilization nation-states o1 o3-mini reasoning-models socioeconomics

Last synced: 17 Mar 2026

https://github.com/contactvaibhavi/hyper-personalised-agent

Hyper-personalised agentic system to aggregate, reason and plan over multiple input streams

agent planning-algorithms reasoning-models retrieval-augmented-generation

Last synced: 11 Jun 2026

https://github.com/akhilpandey95/s1

Experiments on test-time scaling approaches for reasoning LM's to enforce better <think> or <wait> capabilities.

deepseek-r1 inference reasoning-models test-time-computation ttc

Last synced: 03 May 2026

https://github.com/shaheennabi/rlvr_grpo-experiment-with-math500

A small experiment repository comparing a base reasoning model against RLVR-GRPO checkpoints on the Math500 dataset. It includes evaluation results, short-form observations, and a local temp_clone of the full open-posttraining-system codebase for reference.

evaluating-models grpo-checkpoint math500 open-posttraining-system policy-optimization post-training reasoning-models reinforcement-learning rlvr-grpo sparse-rewards

Last synced: 18 Jun 2026

https://github.com/kaicheng001/awesome-r1

A curated list of research papers, models, and resources related to R1-style reasoning models following DeepSeek-R1's breakthrough in January 2025.

awesome deepseek-r1 llm lmm mllm r1 reasoning-models reward-model thinking vlm

Last synced: 22 Jun 2025

https://github.com/angelnicolasc/meridian

Phase-aware vLLM scheduler for reasoning models: output-first dispatch, entropy-gated think termination, tiered KV eviction, and TTOT-focused benchmarking.

cuda inference kv-cache llm observability pyo3 python reasoning-models rust scheduler vllm

Last synced: 27 Jun 2026