Find a story
Search Spins
Search titles, summaries, and missing voices across published articles — press releases, announcements, and media coverage.
7 results for “agentic tasks”
Subagents vs Agent Skills: Executing Reusable Knowledge for Long-Horizon Agentic Tasks
A new arXiv paper proposes 'subagents'—dedicated, context-isolated execution units for agent skills—as a more robust alternative to embedding skill instructions directly into a main agent's context, improving performance on long-horizon tasks where context accumulation degrades reasoning.
Sep 11, 2026
Meta rolls out Muse Spark 1.3 in Muse Code and Meta Model API, saying it significantly improves coding and agentic performance, at the same price as Spark 1.2 (Ina Fried/Axios)
Meta released Muse Spark 1.3, an updated AI model for coding and agentic tasks, claiming significant performance gains without a price increase over version 1.2.
Sep 3, 2026
Attackers Steal METR API Key and Consume AI Credits Worth About $600,000
METR, a nonprofit AI safety evaluator, disclosed two security incidents involving unauthorized access attempts, including theft of an API key that led to $600,000 in unauthorized AI credit consumption.
Sep 1, 2026
Progress-conditioned Group Policy Optimization for Long-Horizon Agentic Tasks
A new reinforcement learning method called ProGPO improves LLM agent training on long-horizon tasks by reweighting credit assignment when all rollouts fail, using state-visit novelty as a proxy for progress.
Jul 28, 2026
Muse Spark 1.1 by Meta AI: Multimodal reasoning model built for agentic tasks - Product Hunt
Meta AI released Muse Spark 1.1, a multimodal reasoning model designed for agentic tasks, as announced on Product Hunt — a platform signaling early user interest but not representing technical validation or deployment evidence.
Jul 11, 2026
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training
Researchers introduced TurnOPD, a turn-aware on-policy distillation method that improves training efficiency and accuracy for long-horizon language agents by reallocating computational budget from low-signal tail turns to deeper decision points.
Jul 9, 2026
Best AI for Agentic Tasks: LLM Leaderboard - Artificial Analysis
An analyst report ranks large language models on 'agentic tasks' using a proprietary benchmark, positioning certain models as leaders in autonomous reasoning and action — but provides no methodology, validation, or independent replication details.
Published Oct 3, 2025 · Analyzed Jul 6, 2026