News / #pricing Tag Pricing changes 410 articles archived under #pricing · RSS Sign in to follow Hacker News — AI on Front Page community 24d ago Beating GPT-5.6 Sol on retrieval with 100x cheaper open models Article URL: https://neon.com/blog/how-castform-neon-beats-frontier-models-on-price-and-efficiency Comments URL: https://news.ycombinator.com/item?id=49186762 Points: 226 # Comments: 47 29 TechCrunch — AI news-outlet 24d ago Hark previews its browser use agent for completing tasks Hark claims that its browser use agent is faster and cheaper than competition. 9 arXiv — Machine Learning research 25d ago Neural Networks with Local Converging Inputs for Efficient Options Pricing Models arXiv:2608.02778v1 Announce Type: new Abstract: We present a novel application of Neural Networks with Local Converging Inputs (NNLCI) to improve the efficiency of existing numerical methods for pricing multi-asset options. The most concise input format for NNLCI has been… 29 Vercel — AI dev-tools 25d ago Full Sandbox egress firewall now available on Hobby plan All Vercel Sandbox firewall features are now available on the Hobby plan. This brings the same network isolation that protects production workloads to the free tier, giving Hobby builders control over exactly what leaves the sandbox while keeping secrets out of the code… 31 r/LocalLLaMA community 25d ago SK hynix, In Collaboration With SanDisk, Unveils The New High Bandwidth Flash (HBF) Standard, Helping To Resolve AI Inference Bottlenecks, Targeting Up To 3TB/s Bandwidth Hopefully this would let us have faster local models....but it will probably be out of our price range.   submitted by   /u/giveen [link]   [comments] 12 arXiv — Machine Learning research 26d ago FedChronos: Federated Fine-Tuning of Time-Series Foundation Models for Privacy-Preserving Commodity Price Forecasting arXiv:2608.01290v1 Announce Type: new Abstract: Time-series foundation models (TSFMs) such as Chronos have demonstrated strong forecasting capabilities across domains, yet adapting them to institutionally fragmented settings, where data cannot be centralized due to regulatory,… 18 arXiv — Machine Learning research 26d ago When May a Model Replace the Experiment? Audits, Licenses, and the Price of Trust in Surrogate-Driven Design arXiv:2608.01378v1 Announce Type: new Abstract: Design campaigns in chemistry, materials science, and machine learning share a bottleneck: determining how good a candidate truly is requires an expensive evaluation - an experiment, a first-principles simulation, or a full… 20 arXiv — NLP / Computation & Language research 26d ago Language Equality has a Price: A Systematic Investigation of Multi-turn LLM Performance for EU-24+ arXiv:2608.01395v1 Announce Type: new Abstract: We evaluate large language models (LLMs) as language agents playing goal-directed dialogue games in self-play across 30 languages: the 24 official EU languages plus six others. Unlike static or preference-based evaluation, this… 5 Dwarkesh Podcast news-outlet 26d ago Why smarter AI models could drive up compute prices 10x The end of cheap compute? 29 r/LocalLLaMA community 27d ago 70-class VRAM stagnation been thinking about how the desktop 70-class has sat at 12GB for two generations now, 4070, 4070 super, 5070, all 12GB. the 1070 gave you 8GB back in 2016 and it felt generous for the price. ten years later the jump is... 4GB. and the thing is these chips arent even weak. the… 32 Smol AI News news-outlet 27d ago Qwen 3.8 Max **Alibaba** launched **Qwen3.8-Max**, a **2.4T-parameter** open-weight model emphasizing autonomous coding, long-horizon execution, and multimodal feedback, with aggressive pricing. Early benchmarks rank it highly on human-preference and vision tasks, showing parity with… 21 arXiv — NLP / Computation & Language research 27d ago Data Turnstile: A Scalable Open Framework for Function-Calling Data Generation arXiv:2607.29250v1 Announce Type: new Abstract: Small language models (SLMs) are attractive for agentic deployment due to low latency, reduced cost, and on-device privacy, yet they struggle with tool-use tasks where training data is scarce and noisy. Unlike larger models, SLMs… 13 r/MachineLearning community 28d ago [D] Self-Promotion Thread Please post your personal projects, startups, product placements, collaboration needs, blogs etc. Please mention the payment and pricing requirements for products and services. Please do not post link shorteners, link aggregator websites , or auto-subscribe links. -- Any abuse… 7 r/LocalLLaMA community 29d ago DeepSeek-V4-Flash-0731: Models you can run locally now have the intelligence score of the top frontier model from March 2026 March 6th, 2026 the highest intelligence index score was 51 for frontier models. deepseek-ai/DeepSeek-V4-Flash-0731 that has an intelligence score of 50. If these benchmarks are accurate, models available to run locally on <8K USD (us prices - just guestimating/not exact)… 19 r/LocalLLaMA community 29d ago Deepseek V4 Flash is now ~#2 open weight model to Kimi K3 and >50x cheaper https://preview.redd.it/h7zv5tb3tmgh1.png?width=2854&format=png&auto=webp&s=507380e8f862c18f10f7c5c84da9e8d1c59139b0 Deepseek's new flash model is unexpectedly cheap and high-performing across useful benchmarks. It's priced at $0.09 / $0.18 per 1M. Truly "intelligence too cheap… 23 r/LocalLLaMA community 29d ago Translation: We had to cut our price by 80% because a open waits model with 284B and 13B active parameter called DeepSeek v4 flash just price/performance mocked us again.   submitted by   /u/InternationalGap3698 [link]   [comments] 38 Hacker News — AI on Front Page community 1mo ago DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis Article URL: https://artificialanalysis.ai/models/deepseek-v4-flash-ga Comments URL: https://news.ycombinator.com/item?id=49120299 Points: 285 # Comments: 139 9 Hacker News — AI on Front Page community 1mo ago DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis Article URL: https://artificialanalysis.ai/models/deepseek-v4-flash Comments URL: https://news.ycombinator.com/item?id=49120299 Points: 423 # Comments: 229 20 Latent.Space news-outlet 1mo ago [AINews] GPT 5.6 price cut by 20%-80%: Cost of GPT 5.4 Intelligence dropped 13x in 4 months due to GPT 5.6 recursive self-optimization Distillation is all you need! 28 arXiv — NLP / Computation & Language research 1mo ago Using Large Language Models for Idea Generation in Innovation arXiv:2607.27553v1 Announce Type: cross Abstract: This research evaluates the efficacy of large language models (LLMs) in generating new product ideas. To do so, we compare three pools of ideas for new products targeted toward college students and priced at 50 dollars or less.… 31 Simon Willison community 1mo ago Advancing the price-performance frontier with GPT‑5.6 Advancing the price-performance frontier with GPT‑5.6 Huge price drop from OpenAI today: GPT-5.6 Terra got a 20% reduction, and GPT-5.6 Luna got a massive 80% drop. OpenAI credit 5.6 Sol with enabling this: in How GPT‑5.6 fuses frontier intelligence with frontier efficiency they… 20 TechCrunch — AI news-outlet 1mo ago Friend, the lonely AI wearable, returns with a new voice and a much bigger price tag Friend, the AI wearable, can now talk to its users — for an enhanced price. 38 Hacker News — AI on Front Page community 1mo ago Advancing the price-performance frontier with GPT‑5.6 Article URL: https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/ Comments URL: https://news.ycombinator.com/item?id=49112867 Points: 262 # Comments: 162 27 r/LocalLLaMA community 1mo ago How close are we to local llama robotics for consumer price point? I'm guessing 3 years, what do you think? In other words: many of us will be able to afford a general purpose robot in 3 years to experiment with in the home. Cost roughly $5k? Probably small size, but hopefully still able to do the dishes and operate a vacuum.   submitted by… 13 OpenAI official-blog 1mo ago Advancing the price-performance frontier with GPT-5.6 Explore lower GPT‑5.6 pricing for Luna and Terra—and how OpenAI’s more efficient models help enterprises deploy AI workflows at scale. 35 Smol AI News news-outlet 1mo ago not much happened today **OpenAI** aggressively cut prices for **GPT-5.6 Luna** by 80% and **Terra** by 20%, introducing a faster **Sol Fast** tier with up to 2.5× lower latency at double the price, improving agent workflow costs by roughly 10×. The **ARC-AGI-3** debate highlighted that the complete… 14 arXiv — Machine Learning research 1mo ago Inverse Learning of Latent Risk-Neutral Densities from Irregular Option Quotes arXiv:2607.27188v1 Announce Type: new Abstract: Accurate option prices do not imply accurate recovery of the latent risk-neutral density. We study this distinction with two complementary benchmarks. A controlled benchmark exposes simulator-truth densities for latent evaluation,… 30 Vercel — AI dev-tools 1mo ago AI Gateway: GPT-5.6 pricing and speed updates On AI Gateway , GPT-5.6 Luna and GPT-5.6 Terra are now cheaper and GPT-5.6 Sol is faster. AI Gateway adds no markup on token pricing, so these changes reach you at the upstream rate. The changes apply to both short and long context pricing. Model Change Input: Short context (per… 4 TechCrunch — AI news-outlet 1mo ago Mark Zuckerberg predicts that billions of people will have personal AI agents in five years As Meta pours billions into AI infrastructure and agents, Zuckerberg is working to convince investors that the payoff will be worth the price. 19 Dwarkesh Podcast news-outlet 1mo ago Why compute might get 10x+ more expensive in coming years If a human-level software engineer that could run on an H100 equivalent, at current market rates for software engineers, that H100 should rent for over $250k a year. That’s 15x today’s spot price. 19 r/LocalLLaMA community 1mo ago dropped 4k on a spark, am I crazy? Saw that the Asus Ascent 1tb was going for $3,950 from a few sources, couldn't stop thinking about it, finally just went ahead and did it. Am I completely insane? Will I regret this? I can't imagine the price will go down any time soon, so it seems like a good idea and I… 34 r/LocalLLaMA community 1mo ago Nvidia is expected to raise GeForce RTX GPU prices again by up to 30%   submitted by   /u/ab2377 [link]   [comments] 21 r/LocalLLaMA community 1mo ago I've been tracking RTX 5090 prices across EU stores since March, it's up €1,061 and still climbing Been running a GPU price tracker ( https://www.pricesquirrel.com ) since March, covering 20+ EU stores, recently added RAM, SSDs and CPUs too. Every GPU tier has gotten cheaper since launch. The RTX 5090 has done the exact opposite. The data: The ASUS TUF Gaming RTX 5090 OC was… 34 TechCrunch — AI news-outlet 1mo ago Cursor makes its biggest India push yet ahead of SpaceX acquisition with localized pricing Cursor says India is now its third-largest market globally and plans to expand local hiring and enterprise sales. 23 arXiv — Machine Learning research 1mo ago Optimizing Transformer Neural Network for Real-Time Outlier Detection on FPGAs arXiv:2607.22786v1 Announce Type: new Abstract: In this work, we explore how the inference time of a Transformer Neural Network can be efficiently optimized with applications to real-time anomaly detection in financial time series. The financial time series are price series such… 38 arXiv — Machine Learning research 1mo ago Bitcoin Price Direction Prediction via Regime-Aware Multi-Modal Fusion of Social Sentiment and Technical Features arXiv:2607.23370v1 Announce Type: new Abstract: Bitcoin price prediction on sub-daily timescales is a hard open problem in computational finance. Bitcoin exhibits fat-tailed returns, non-stationary dynamics, and a price discovery process influenced by social discourse on Reddit… 23 r/LocalLLaMA community 1mo ago First evidence of a pending qwen3.7 open weights release. Qwen3.7-flash is on open router. They referred to Qwen3.6-35b-a3b as Qwen3.6 flash so this is likely a small MoE. The prices are substantially cheaper than 3.6 flash with a native 1M context window.   submitted by   /u/fulgencio_batista [link]   [comments] 37 r/LocalLLaMA community 1mo ago I want to run Kimi K3 at home, so I’m trying to make 2.8T-scale experimentation cheaper Hey r/LocalLLaMA , I’m a retired engineer with a background in distributed computing, currently running a 1-person startup. Like many people here, I’d love to experiment with 2T+ MoE models locally. The problem is that I don’t have an H100 cluster in my living room. So I’ve been… 12 Smol AI News news-outlet 1mo ago not much happened today **Alibaba** launched **Qwen3.8-Max**, a **2.4T-parameter** open-weight model emphasizing autonomous coding, long-horizon execution, and multimodal feedback, with aggressive pricing. Early benchmarks rank it highly on human-preference and vision tasks, showing parity with… 33 r/LocalLLaMA community 1mo ago Will prices finally go down? I am seeing more and more videos as posts about how OpenAI is in complete financial ruin, Anthropic isn't much better. Their expenses go with the revenue they make etc etc. Meta made big investments into AI data centers and had no use for then, had to rent them, same thing with… 38 r/LocalLLaMA community 1mo ago Do people building local LLM rigs track RTX Ada/workstation card prices, or just consumer cards like the 5090? curious how people here approach buying high-end/workstation cards (RTX 6000 Ada, 5000 Ada, etc) for local LLM work, do you actively watch pricing/timing on these specifically, or is the consumer 5090 usually enough for most builds? also wondering if price alerts/tracking tools… 15 Simon Willison community 1mo ago An Inside Look at the Relay Market Powering Token Resellers and Fraud An Inside Look at the Relay Market Powering Token Resellers and Fraud Fascinating investigation by Matt Lenhard into the market that has grown up around reselling LLM tokens at a discount by pooling API keys from various sources. This looks to be mostly a thing in China.… 18 r/LocalLLaMA community 1mo ago Any use cases for RTX PRO 4500? At its price point, PRO 4500 doesn’t offer as much raw performance due to its lower power draw at 300W. The 5090 can perform up to 60-70% in short spurts with 600W, but can also be undervolted down to 400W. Are there legitimate reasons other than 24/7 usage and lower power draw… 27 r/LocalLLaMA community 1mo ago Is it worth getting 128GB MacBook Pro? Will it ever be comparable to today’s frontier models for coding? I am a long time iOS app developer. In the last year I have been using Cursor+Claude/others to assist with app development. I am concerned that the current low pricing will disappear eventually. I am pricing out a new laptop with the intention of using local models instead. New… 35 Latent.Space news-outlet 1mo ago [AINews] Claude Opus 5: Fable-level performance at Opus price (half Fable) ain't nobody beats Anthropic at distilling Fable! 19 Ars Technica — AI news-outlet 1mo ago Anthropic's Opus 5 is about token efficiency, not a capability leap Models are improving quickly, but the cheaper options are often good enough. 28 TechCrunch — AI news-outlet 1mo ago Anthropic launches Opus 5 Opus 5 will be both cheaper and less restrictive than Fable, likely making it preferable in most use cases 27 r/MachineLearning community 1mo ago I built an open-source multi-agent SDLC harness that beats a cold Claude Code run on large repos, by learning the repo once. Real benchmarks (incl. where it loses) inside. [P] Built an open-source AI coding agent that was 7%–75% cheaper than a cold "claude -p" run on 6/6 well-localized tasks across repositories up to ~82k LOC. The biggest difference: Cold agent: $6.83, 207 turns AutoDev Studio: ~$1.70 for the same bug The full benchmark (including… 14 arXiv — Machine Learning research 1mo ago Adaptive Multi-Horizon Reinforcement Learning arXiv:2607.20656v1 Announce Type: new Abstract: Effective decision-making in complex and changing environments requires balancing short-term and long-term consequences. In reinforcement learning (RL), this trade-off is typically controlled through a fixed discount factor, which… 30 r/LocalLLaMA community 1mo ago How much are RTX PRO 6000s going for in your country/state? Hello guys, hoping you're doing fine! On the last 2-3 months, price of the RTX 6000 PRO seem to have gone insane. I will start on the price here on my country, Chile: RTX 6000 PRO Workstation Edition: 21382 USD post 19% tax. RTX 6000 PRO MaxQ Workstation Edition: 20669 USD post… 29 Page 3 of 9 · 410 articles ← Newer Older →