Skip to content
Read the original: Meituan LongCat· Meituan_LongCat·Published· Aug 28, 2026AI score62

Meituan LongCat Study Tests Whether AI Agents Can Do Research

AI agents can now propose changes, run experiments, interpret feedback, and refine technical artifacts over many iterations.

AISummary

Meituan LongCat evaluated 7 frontier models on 36 AI R&D tasks covering 756 trajectories, looking beyond final scores. Of 252 solutions, only 3 were novel approaches, and most adapted or combined established techniques. The authors conclude that current agents work more like engineering optimizers than autonomous researchers, with reliability, experience reuse, and novelty still open challenges.

Read the original x.com

Source: Meituan LongCat · x.com