All posts

reinforcement learning

4 posts

Cognition Launches SWE-1.7: A 1T Parameter MoE Coding AI That Thinks Before It Acts (1000 TPS)
AgentsJul 8, 2026·4 min read

Cognition Launches SWE-1.7: A 1T Parameter MoE Coding AI That Thinks Before It Acts (1000 TPS)

Cognition releases SWE-1.7, detailing its distributed multi-cluster reinforcement learning architecture, entropy preservation, and context management.

Read more →
DeepReinforce Launches Ornith-1.0: Self-Scaffolding LLMs for Agentic Coding
AnnouncementsJun 25, 2026·5 min read

DeepReinforce Launches Ornith-1.0: Self-Scaffolding LLMs for Agentic Coding

DeepReinforce has released Ornith-1.0, a self-improving family of open-source models utilizing a novel self-scaffolding RL framework for agentic coding.

Read more →
AstraFlow: Infini-AI-Lab's Open-Source Dataflow RL System for Multi-Agentic LLM Training
AI ResearchMay 21, 2026·4 min read

AstraFlow: Infini-AI-Lab's Open-Source Dataflow RL System for Multi-Agentic LLM Training

Infini-AI-Lab has released AstraFlow, an open-source dataflow-oriented reinforcement learning system built specifically for training multi-agentic and multi-policy LLMs. It achieves 2.7x faster multi-policy collaborative RL training, reduces remote rollout sync from 28 GB to 1.5 GB, and supports elastic deployment across heterogeneous GPUs with zero-code configuration.

Read more →
Prime Intellect Introduces Renderers: 3x Throughput for Agentic RL Training
AI ResearchMay 16, 2026·4 min read

Prime Intellect Introduces Renderers: 3x Throughput for Agentic RL Training

Prime Intellect's Renderers fix the token-message mismatch between RL trainers and agent environments, unlocking more than 3x throughput improvement on popular open models without changing the underlying architecture.

Read more →

🚀 Join the AI dev community — follow us everywhere

© 2026 MARKTECHPOST AI MEDIA INC. All rights reserved.Terms & ConditionsPrivacy Policy
Beta Mode