Blogs

Reward hacking is swamping model intelligence gains · Cursor

Cursor

On SWE-bench Pro, 63% of successful Opus 4.8 Max resolutions retrieved the fix rather than derived it.

Visit Site

Blogs Cursor

BlogsHow Wayfair cut ML model costs by 90% (twice!) with Cursor · CursorCursor BlogsGrok 4.5 Model Card · CursorCursor BlogsHow Cursor Router chooses the right model for the task · CursorCursor BlogsAgent swarms and the new model economics · CursorCursor BlogsCoinbase reduces time from idea to production by 90% with Cursor · CursorCursor BlogsIntroducing Cursor Start · CursorCursor ResearchA Guide to Extended Threat Detection and Response: What It Is and How to Choose the Best SolutionsCybersecurity Exchange BlogsClickStack and Hud bring runtime intelligence to AI-powered developmentClickhouse BlogsTrain past the frontier: Training API now generally availableFireworks BlogsDeepSeek V3 just got vision capabilities!Fireworks BlogsVision Model Platform Updates: Enhanced Capabilities and New FeaturesFireworks BlogsIntroducing FireRouter with OpusFireworks BlogsIntroducing Fireworks on Microsoft Foundry: Bringing Best-in-Class Open Model inference to AzureFireworks BlogsAccelerate your Vision Pipelines with the new NVIDIA Nemotron Nano 2 VL Model on FireworksFireworks BlogsIntroducing Supervised Fine-tuning V2Fireworks BlogsFireworks Nexus: Drop-in Open Frontier Intelligence for Teams with BudgetsFireworks BlogsBeyond Supervised Fine Tuning: How Reinforcement Learning Empowers AI with Minimal LabelsFireworks BlogsThe DeepSeek Model Lineup: V3.2, R1, and Distilled Variants Mapped to Production WorkloadsFireworks BlogsBuild a Talking Halloween Skeleton with Docker Model RunnerDocker BlogsWhat Is Code?Martinfowler