Cheaper, Better, Faster, Stronger
Related stories
When to Use Which? Benchmarking Optimisers for Configurable Systems under Varying Budgets
arXiv:2607. 16476v1 Announce Type: cross Abstract: Software configuration tuning is crucial for optimising system performance, and various optimisers have emerged over the last decade.
Introducing our new pricing
Why Specialization Is Inevitable
Brevity is the Soul of Inference Efficiency: Inducing Concision in VLMs via Data Curation
Inference efficiency is typically pursued by shrinking the model: distillation, pruning, quantization, and sparse routing each lower per-token cost while treating token count as fixed. But output length has been inflating, and it is precisely the component the standard toolkit leaves untouched.
How Much Due Diligence Before You Bid? Learning in Intractable Takeover Auctions
arXiv:2606. 29457v1 Announce Type: new Abstract: When two companies bid to buy the same target, no one knows exactly what the target is worth.
Brevity is the Soul of Inference Efficiency: Inducing Concision in VLMs via Data Curation
arXiv:2606. 25432v1 Announce Type: new Abstract: Inference efficiency is typically pursued by shrinking the model: distillation, pruning, quantization, and sparse routing each lower per-token cost while treating token count as fixed.
Brick: Spatial Capability Routing for the Mixture-of-Models (MoM) Paradigm
arXiv:2606. 13241v1 Announce Type: new Abstract: Defining query difficulty is one of the hardest problems in deployment engineering.
Memory Scarcity, Open Models, and the Restructuring of the AI Industry, 2026-2030 -- A quantitative scenario analysis of inference economics, training-cost divergence, and infrastructure solvency
arXiv:2607. 07207v1 Announce Type: cross Abstract: We analyze how four forces restructure the AI industry over 2026-2030: the DRAM/HBM price surge, frontier-capable open-weight models (GLM-5.
A Theory of Training Profit-Optimal LLMs
arXiv:2605. 16430v2 Announce Type: replace-cross Abstract: Scaling LLMs requires tremendous computational resources, and recent advances in AI have gone hand in hand with massive amounts of capital expenditure.
Code Is Cheap. Engineering Judgement Is Now the Scarce Resource
The barriers to building have collapsed. That shifts the bottleneck to ownership, validation, taste, and deciding what should actually exist The post Code Is Cheap.
The Price of Intelligence: A Quality-Adjusted Price Index for AI Services
arXiv:2608.29843v1 Announce Type: cross Abstract: Posted prices for AI inference have fallen steadily since 2024, yet the measured speed of that fall depends almost entirely on the method of measurem...