
Cognition Launches SWE-1.7: A 1T Parameter MoE Coding AI That Thinks Before It Acts (1000 TPS)
Cognition releases SWE-1.7, detailing its distributed multi-cluster reinforcement learning architecture, entropy preservation, and context management.
Read more →
DeepReinforce Launches Ornith-1.0: Self-Scaffolding LLMs for Agentic Coding
DeepReinforce has released Ornith-1.0, a self-improving family of open-source models utilizing a novel self-scaffolding RL framework for agentic coding.
Read more →
AstraFlow: Infini-AI-Lab's Open-Source Dataflow RL System for Multi-Agentic LLM Training
Infini-AI-Lab has released AstraFlow, an open-source dataflow-oriented reinforcement learning system built specifically for training multi-agentic and multi-policy LLMs. It achieves 2.7x faster multi-policy collaborative RL training, reduces remote rollout sync from 28 GB to 1.5 GB, and supports elastic deployment across heterogeneous GPUs with zero-code configuration.
Read more →
Prime Intellect Introduces Renderers: 3x Throughput for Agentic RL Training
Prime Intellect's Renderers fix the token-message mismatch between RL trainers and agent environments, unlocking more than 3x throughput improvement on popular open models without changing the underlying architecture.
Read more →