databricks switched from opus to glm, which their evals found had 92% of the pass rate at 1/4th the cost. ...their evals on their own proprietary codebases, not on public benchmarks that models could have hillclimbed. the chinese benchmaxxing accusations are officially cope.
Today
@databrickswe're publishing a detailed analysis of techniques we used to drastically reduce our internal AI spend while aggressively growing adoption. Savings come from layering in several techniques, which combine to drive unit costs down as much as 90% in some scenarios.
Incredibly grateful to stand with all the leaders in this space rallying around open weights. Thank you to everyone that has contributed to open source, whether that's a PR, a research paper, or a signature.
cloud was effectively commoditized in march 2014 when google cut compute prices 32%, and aws was forced to match within days with its 42nd price cut at the time. inference is speedrunning the same story, when dollars are involved markets are brutally efficient.



