Goblin
News
AI news by
promptgoblins.ai
|
News
About
News
About
Filtered by:
Cost Optimization
Clear
Titles
Summaries
6
Pi Coding Agent Cuts Context Costs 88% with Disk-Spill Architecture
Infra
1
Aug 25
6
Pi Coding Agent Cuts Context Costs 88% with Disk-Spill Architecture
Infra
· 1 src · Aug 25
Discuss
7
AT&T Targets 70–80% Open Model AI Usage, Processes 45 Billion Tokens Per Day
Updated Aug 20
Enterprise
2
Aug 20
7
AT&T Targets 70–80% Open Model AI Usage, Processes 45 Billion Tokens Per Day
Upd Aug 20
Enterprise
· 2 srcs · Aug 20
Discuss
6
NVIDIA NeMo Switchyard Benchmark: Only 7% of Agent Calls Need a Frontier Model, Cutting Costs 74%
Research
1
Aug 12
6
NVIDIA NeMo Switchyard Benchmark: Only 7% of Agent Calls Need a Frontier Model, Cutting Costs 74%
Research
· 1 src · Aug 12
Discuss
6
Databricks Achieves 70% AI Coding Cost Reduction, Publishes Enterprise Playbook with Stripe, Coinbase, and Uber Insights
Enterprise
1
Aug 7
6
Databricks Achieves 70% AI Coding Cost Reduction, Publishes Enterprise Playbook with Stripe, Coinbase, and Uber Insights
Enterprise
· 1 src · Aug 7
Discuss
6
Research: Structured Distillation Cuts AI Agent Memory Tokens 11× While Preserving Retrieval Accuracy
Research
1
Aug 6
6
Research: Structured Distillation Cuts AI Agent Memory Tokens 11× While Preserving Retrieval Accuracy
Research
· 1 src · Aug 6
Discuss
6
Cursor Launches Router: Intelligent Model Routing Cuts AI Coding Costs by 60%
Products
1
Jul 23
6
Cursor Launches Router: Intelligent Model Routing Cuts AI Coding Costs by 60%
Products
· 1 src · Jul 23
Discuss
6
Harness Tuning Alone Brings Nemotron 3 Ultra to Near-Frontier Agent Performance at 10x Lower Cost Than Claude Opus 4.8
Infra
1
Jul 8
6
Harness Tuning Alone Brings Nemotron 3 Ultra to Near-Frontier Agent Performance at 10x Lower Cost Than Claude Opus 4.8
Infra
· 1 src · Jul 8
Discuss
7
Cognition Launches Devin Fusion: Frontier-Quality Coding Agent at 35% Lower Cost
Products
1
Jun 30
7
Cognition Launches Devin Fusion: Frontier-Quality Coding Agent at 35% Lower Cost
Products
· 1 src · Jun 30
Discuss
6
AWS Two-Model Pipeline Pairs Nova 2 Lite with Claude Sonnet 4.6 for 33% Cheaper Document Processing
Products
1
Jun 29
6
AWS Two-Model Pipeline Pairs Nova 2 Lite with Claude Sonnet 4.6 for 33% Cheaper Document Processing
Products
· 1 src · Jun 29
Discuss
6
Wayfinder Router Offers Deterministic, Offline LLM Query Routing Without Model Calls
Open Source
1
Jun 28
6
Wayfinder Router Offers Deterministic, Offline LLM Query Routing Without Model Calls
Open Source
· 1 src · Jun 28
Discuss
6
Fireworks AI CEO Lin Qiao: Per-Application Custom Models Are the Key to Sustainable AI Agent Economics
Infra
1
Jun 18
6
Fireworks AI CEO Lin Qiao: Per-Application Custom Models Are the Key to Sustainable AI Agent Economics
Infra
· 1 src · Jun 18
Discuss
6
LangChain and Fireworks Build 100x Cheaper AI Trace Judge Using Fine-Tuned Qwen Model
Products
1
Jun 15
6
LangChain and Fireworks Build 100x Cheaper AI Trace Judge Using Fine-Tuned Qwen Model
Products
· 1 src · Jun 15
Discuss
Filters
Signal
Title
Category
Sources
Posted
Discuss