rohitg00/ai-engineering-from-scratch
Learn it. Build it. Ship it for others.
Explore 19 GitHub repositories focused on reinforcement-learning. Discover top-starred projects and those trending this week.
Learn it. Build it. Ship it for others.
[Up-to-date] A curated list of resources on graph-empowered agents and agent-facilitated graph learning (Graphs Meet Agents & Agentic Graph Engineering).
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, FLUX and more.
《深入理解 AI Agent:设计原理与工程实践》(李博杰 著)开源主仓库:全书正文、编译版 PDF 与按章配套代码
The absolute trainer to light up AI agents.
Infrastructure for continually self‑improving agents
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
OpenDILab Decision AI Engine. The Most Comprehensive Reinforcement Learning Framework B.P.
A modular, primitive-first, python-first PyTorch library for Reinforcement Learning.
🦞 Just talk to your agent — it learns and EVOLVES 🧬.
[NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments
Isaac Lab API, powered by MuJoCo-Warp, for RL and robotics research
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, FLUX and more.
Learn it. Build it. Ship it for others.
《深入理解 AI Agent:设计原理与工程实践》(李博杰 著)开源主仓库:全书正文、编译版 PDF 与按章配套代码
The absolute trainer to light up AI agents.
Infrastructure for continually self‑improving agents
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
OpenDILab Decision AI Engine. The Most Comprehensive Reinforcement Learning Framework B.P.
A modular, primitive-first, python-first PyTorch library for Reinforcement Learning.
🦞 Just talk to your agent — it learns and EVOLVES 🧬.
[NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments
Isaac Lab API, powered by MuJoCo-Warp, for RL and robotics research
Awesome Reasoning LLM Tutorial/Survey/Guide
Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch
🐝 The First Self-Improving agents with RL / Prompting Optimization
UniRL is a Framework for Unified Multimodal Model Reinforcement Learning
A Python library for running thousands of MuJoCo simulations in parallel on CPU
[Up-to-date] A curated list of resources on graph-empowered agents and agent-facilitated graph learning (Graphs Meet Agents & Agentic Graph Engineering).
Open skill for capturing AI agent work as structured traces.
Multi-agent simulation library in Python