Atropos
NousResearch/atropos · Published by Nous Research
Archived Nous Research framework of RL environments for collecting and evaluating LLM trajectories
Atropos is an archived Nous Research environment microservice framework for asynchronous reinforcement learning with language models. Its README links DeepHermes model artifacts trained with its environments.
What Atropos does
Atropos is Nous Research's environment microservice framework for asynchronous reinforcement learning with language models. Environments run as services and send data to a trajectory API, from which a trainer pulls batches. The trainer and inference engine are not included in the package. The framework supports dataset environments such as GSM8K and MMLU, online game environments such as Blackjack and Taxi, RLAIF and RLHF, multi-turn RL, code execution and multimodal tasks.
The README reports results from models trained with its environments, such as a tool-calling specialist whose Berkeley Function Calling score on parallel tasks rose from 10% to 46%, and a financial-fundamentals model whose directional prediction accuracy rose from 20% to 50%. It links DeepHermes model artifacts on Hugging Face, including RLAIF personality experiments. The repository is archived and no longer maintained.
Key features
- Environment microservices with a trajectory API for trainers to pull batches
- Dataset environments such as GSM8K, MMLU and custom Hugging Face datasets
- Online game environments including Blackjack, Taxi and text-based games
- RLAIF, RLHF and multi-turn RL environments
- Code execution (MBPP, HumanEval) and multimodal (OCR VQA, Clevr) environments
When to use it
- Collect LLM trajectories from interactive environments for RL training
- Evaluate a model on dataset or game environments
- Study how the DeepHermes artifacts linked in the README were produced
Who it is for: Researchers studying reinforcement learning environments for language models.
How it fits with Hermes Agent
A Nous Research repository whose README links DeepHermes models trained with its environments; it concerns training and evaluating LLMs rather than running Hermes Agent.
Note: The repository is archived and no longer maintained, with no further updates, bug fixes or security patches.
FAQ
What is Atropos?
Atropos is Nous Research's environment microservice framework for asynchronous reinforcement learning with language models. Environments feed trajectories to a trajectory API that a trainer pulls from.
Does Atropos work with Hermes Agent?
Atropos is a training framework, not an agent runtime. Its README links DeepHermes models trained with its environments but does not describe running Hermes Agent with it.
Is Atropos free and open source?
Yes, the repository is released under the MIT license. It is archived, so no further updates, bug fixes or security patches will be made.
Similar official for Hermes Agent
All officialOptimizes Hermes Agent skills with DSPy and GEPA through API calls, with no GPU training required
NousResearch Hermes Paperclip AdapterAdapter that runs Hermes Agent as a managed employee inside a Paperclip company
NousResearch autonovelAutonomous pipeline that writes, revises, typesets, illustrates and narrates a complete novel
NousResearch Hermes Function CallingReference code for Hermes Pro models to run function calling and JSON mode against a tool schema
NousResearch Hermes Bot ModeArchived plugin that turns Hermes profiles into named bots; now bundled in Hermes Desktop
NousResearch AutoreasonPaper and experiment code for a self-refinement method that treats doing nothing as a valid option
Related guides: What is Hermes Agent? · How to install Hermes Agent