Goblin
News
AI news by
promptgoblins.ai
|
News
About
News
About
Filtered by:
SWE-Bench
Clear
Titles
Summaries
7
LLM-as-a-Verifier Framework Beats SOTA on Coding Benchmarks While Cutting Costs 11×
Research
1
Aug 19
7
LLM-as-a-Verifier Framework Beats SOTA on Coding Benchmarks While Cutting Costs 11×
Research
· 1 src · Aug 19
Discuss
7
Cognition Launches SWE-1.7, Claiming Frontier Coding Performance at Lower Cost via Cerebras
Research
1
Jul 9
7
Cognition Launches SWE-1.7, Claiming Frontier Coding Performance at Lower Cost via Cerebras
Research
· 1 src · Jul 9
Discuss
8
DeepReinforce Open-Sources Ornith-1.0: Self-Scaffolding Agentic Coding Models from 9B to 397B Parameters
Open Source
1
Jun 26
8
DeepReinforce Open-Sources Ornith-1.0: Self-Scaffolding Agentic Coding Models from 9B to 397B Parameters
Open Source
· 1 src · Jun 26
Discuss
Filters
Signal
Title
Category
Sources
Posted
Discuss