News / #pricing Tag Pricing changes 410 articles archived under #pricing · RSS Sign in to follow r/LocalLLaMA community 1mo ago Now brothers we know why we are so fucked up Samsung chip division's single-year profits beat its past 40 years of profits, combined, due to increased memory and storage prices — Samsung passes Nvidia to become most profitable company in the world, notches 19x quarterly increase in profit… 33 r/MachineLearning community 1mo ago Why does the same H100 cost 5x more depending on where you rent it? [D] I kept finding wildly different prices for the same GPU across providers and data centers, so I built a OS CLI that searches live GPU capacity and shows the cheapest available routes npx gpu-price-finder Supports RTX 4090, RTX 5090, L40S, A100, H100 and lets you filter by… 20 TechCrunch — AI news-outlet 1mo ago SpaceXAI releases Grok 4.5, which Elon describes as an ‘Opus-class model’ Elon Musk's tech company released the newest version of Grok on Wednesday, promising a cheaper, more efficient alternative to other powerful AI models. 30 r/LocalLLaMA community 1mo ago Mimo & deepseek are really amazing at optimizing ai. Read the the official blog page i linked, it will give amazing insight on how they pulled off this kind of low pricing with 2x - 3x profit margins. For quick look --> https://x.com/i/status/2059618247553745204 Detailed --> https://mimo.xiaomi.com/blog/mimo-v2-5-inference I hope in future we get fable lvl ai at the cost of current DSV4. Thats far more sufficient for like 90% of people. Xai is also pushing for low cost api… 17 arXiv — Machine Learning research 1mo ago Evaluating Time Series Foundation Models for Electricity Price Forecasting: Contamination Risk, Distributional Shifts, and Covariate Dependence arXiv:2607.02623v1 Announce Type: new Abstract: Time series foundation models (TSFMs) have shown strong zero-shot forecasting performance, but their generalization in covariate-driven, non-stationary settings is underexplored. Electricity price forecasting (EPF) presents a… 12 arXiv — Machine Learning research 1mo ago Trading Confidence: Comprehensive Uncertainty Estimation in Algorithmic Trading arXiv:2607.02864v1 Announce Type: new Abstract: Reinforcement Learning (RL) has emerged as a powerful approach in financial trading, enabling agents to learn optimal strategies through direct market interaction. However, financial markets are highly uncertain, with price… 12 arXiv — Machine Learning research 1mo ago Dynamic Regret for Non-Stationary Linear Bandits via Misspecification Reductions arXiv:2607.02891v1 Announce Type: new Abstract: Many online decision-making problems involve both round-specific feasible actions and drifting reward models: eligible ad impressions, feasible prices, and available treatments can change over time, while user preferences, demand… 34 r/LocalLLaMA community 1mo ago Blower 5060 Ti's from Alibaba, good idea? Asking price is 580 USD on Alibaba, maybe stacking a few of these together for a "power-efficient" system is a good idea? They have 8 pin connectors as usual. I'm aware the 3080 20GB exists for 500 USD with 25% more bandwidth, but they require a larger PSU and run hotter also… 37 TechCrunch — AI news-outlet 1mo ago Vercel CEO Guillermo Rauch on the fight to split off models from agents "The reality is, when you're optimizing for production, you start looking at a price/performance," Guillermo Rauch tells TechCrunch. 34 Hacker News — AI on Front Page community 1mo ago Performance per dollar is getting faster and cheaper Article URL: https://www.wafer.ai/blog/glm52-amd Comments URL: https://news.ycombinator.com/item?id=48780417 Points: 218 # Comments: 68 11 Hacker News — AI on Front Page community 1mo ago The Egg Bandits Made a Thousand Times the Fine They Just Paid for Price Fixing Article URL: https://www.thebignewsletter.com/p/crime-pays-the-egg-bandits-made-a Comments URL: https://news.ycombinator.com/item?id=48761229 Points: 230 # Comments: 97 6 arXiv — Machine Learning research 1mo ago Neural Certificate Pricing for Combinatorial Optimization Problems arXiv:2607.01185v1 Announce Type: new Abstract: Combinatorial optimization (CO) problems are difficult because certifiable discrete structure induces exponential search. One needs to search over the set exponentially many candidates to certify optimality, however, the structural… 23 r/MachineLearning community 1mo ago [D] Self-Promotion Thread Please post your personal projects, startups, product placements, collaboration needs, blogs etc. Please mention the payment and pricing requirements for products and services. Please do not post link shorteners, link aggregator websites , or auto-subscribe links. -- Any abuse… 17 r/MachineLearning community 1mo ago Spot/interruptible H100 and A100 pricing across RunPod, Vast.ai, and AWS - June 2026 data [D] Following up on the on-demand comparison from a couple weeks back - pulled spot/ interruptible pricing this time since that's where the real savings conversation actually lives for anyone running checkpointed training or batch jobs. Checked: June 2026. Spot/interruptible tier,… 35 TechCrunch — AI news-outlet 2mo ago Google introduces a faster, cheaper image generator with Nano Banana 2 Lite Google is updating its image generator to make it faster and cheaper, making it a more useful tool for creators looking to make AI content. 15 TechCrunch — AI news-outlet 2mo ago Anthropic launches Claude Sonnet 5 as a cheaper way to run agents Anthropic’s Claude Sonnet 5 brings stronger agentic capabilities, lower pricing, and improved safety, positioning the model as a cheaper alternative to Opus, GPT-5.5, and Gemini Pro. 5 r/LocalLLaMA community 2mo ago Tesla V100 16GB local LLMs, single and dual NVLink benchmarks Picked up a couple of Tesla V100-SXM2-16GB modules a while back to run local models and drive Claude Code fully offline, figured the actual numbers and the traps might save someone else the pain. They've come right down in price and the 16GB of HBM2 at ~900 GB/s still holds up… 33 arXiv — Machine Learning research 2mo ago When Prices Double in a Week: Forecasting of Agricultural Volatility in Import-Isolated Markets arXiv:2606.29248v1 Announce Type: new Abstract: Vegetable prices in Sri Lanka are highly volatile because the market is largely import-isolated, so supply disruptions quickly drive prices up. This study develops a machine learning framework to forecast such volatility by… 8 Vercel — AI dev-tools 2mo ago Claude Sonnet 5 now available on Vercel AI Gateway Claude Sonnet 5 from Anthropic is now available on AI Gateway . Sonnet 5 improves on Sonnet 4.6 across coding and agentic work, reaching outcomes on many tasks that previously needed an Opus model, at Sonnet pricing. The model is more agentic and follows instructions more… 14 Vercel — AI dev-tools 2mo ago Vercel Agent has updated pricing Vercel Agent pricing is changing. You no longer need to pre-load credits or manage a separate wallet for Vercel Agent. Instead of a $0.30 per-request fee, you now pay a Vercel Token Rate of $0.25 per million tokens, plus provider inference costs at the underlying token rate. The… 6 r/MachineLearning community 2mo ago Price elasticity model [R] Need to build a ml model to find the price elasticity at the product group level first the given price and discount. What are features I need have and which model used in the industry for these type of use cases . I have used regression and random regression to predict the qty… 30 TechCrunch — AI news-outlet 2mo ago Anthropic and Gov. Newsom forge deal allowing California government to use Claude at half price As Anthropic forges a closer relationship with the state of California, the federal government has made an enemy out of the OpenAI rival. 26 r/LocalLLaMA community 2mo ago Samsung, SK hynix, Micron Sued in US Over Memory Price Fixing   submitted by   /u/johnnyApplePRNG [link]   [comments] 15 Hacker News — AI on Front Page community 2mo ago Samsung, SK Hynix, Micron Sued in US over Memory Price Fixing Article URL: https://en.sedaily.com/international/2026/06/29/samsung-sk-hynix-micron-sued-in-us-over-memory-price-fixing Comments URL: https://news.ycombinator.com/item?id=48718102 Points: 225 # Comments: 108 16 r/LocalLLaMA community 2mo ago Deepseek V4 Official Launch to be released mid-July with API price changes Is this the official release for deepseek? I hope it has huge improvements https://preview.redd.it/dm5l0qn8k7ah1.png?width=694&format=png&auto=webp&s=12eadfd0a52c0f1a65bcd685f2cdbb29aff457be   submitted by   /u/jmorant555 [link]   [comments] 22 r/LocalLLaMA community 2mo ago AMD MI210 64GB vs DCU K100 64GB On the Chinese eBay there is a many DCU K100 64 GB GPU available for a very attractive price, between 6000 RMB and 19 000 (air or water cooled versions, new or second hands), and 15 000 to 20 000 for the AMD MI210 (4000-6000 RMB for the PCIE bridge). There is very little… 25 arXiv — NLP / Computation & Language research 2mo ago Supersede: Diagnosing and Training the Memory-Update Gap in LLM Agents arXiv:2606.27472v1 Announce Type: new Abstract: Large language model (LLM) agents operate over long, multi-session interactions in which facts change: a user moves, a price updates, a plan is revised. Acting correctly requires using the current value of a fact and discarding… 16 Hacker News — AI on Front Page community 2mo ago Historical memory prices 1960-2026 Article URL: https://dam.stanford.edu/memory-prices.html Comments URL: https://news.ycombinator.com/item?id=48710092 Points: 253 # Comments: 91 10 r/LocalLLaMA community 2mo ago A lot of good M5 Max options available at Apple Refurbished Just a heads-up. After Apple's price hike announcement, they added a bunch of top-of-the-line 14" M5 Pro/Max options to their refurbished website. If you got discouraged by the price hike, check out their refurbished store.   submitted by   /u/Hanthunius [link]  … 13 r/LocalLLaMA community 2mo ago Sell ddr5 for vram? Hi, I have 768gb ddr5 6400 ecc ram, given current ram prices should I sell half and buy rtx 6000 pros? Edit: currently have an epyc 9255 and 2 x rtx 6000 pro max q but could add another three Blackwell potentially then and run something like glm 5.2 in q4 at faster speeds then… 6 r/LocalLLaMA community 2mo ago Finally.. my rig is maxed out Got all the parts before the crazy price increase except for the rtx pro 5k! Was saving up to order rtx pro 6000 in US and i did, but wanted to join nvidia inception program for the discount. It was around $8.5k during that time, less 1k if I succeeded. It took around 3 months… 7 r/LocalLLaMA community 2mo ago How to distill my own models? I've been using cloud provided models for agentic theorem proving a lot, and cost is becoming an issue for me. I have funding for hardware cost but I can't use them for LLM credits which put me in a unique situation where it might be cheaper to self-host models instead of paying… 29 r/LocalLLaMA community 2mo ago Took the plunge! (Minisforum MS-S1 Max) With Apple prices entering the stratosphere, the recent Fable gov't rug pull, and the inevitable closed-model price increases, I decided to pick up a (lightly) used Minisforum MS-S1 Max with 128GB of memory. Comes with a 10-day return and a 3-month warranty. Paid the local equiv… 30 Simon Willison community 2mo ago Quoting OpenAI We're beginning a limited preview of the GPT‑5.6 series: Sol, our flagship model; Terra, a balanced model for everyday work; and Luna, a fast and affordable model. Terra has competitive performance to GPT‑5.5 while being 2x cheaper and Luna brings strong capability at our lowest… 34 TechCrunch — AI news-outlet 2mo ago Early Bird pricing ends tonight for TechCrunch Founder Summit Save up to $190 on your pass to TechCrunch Founder Summit 2026. Early Bird pricing ends today, at 11:59 p.m. PT, after which rates increase. Register now. 11 arXiv — Machine Learning research 2mo ago Embedding Foundation Model Predictions in Discrete-Choice Models with Structural Guarantees arXiv:2606.26432v1 Announce Type: new Abstract: Tabular foundation models achieve strong accuracy on choice prediction tasks, but their predictions often violate the economic logic those tasks require: raising a price can increase predicted demand, implied willingness-to-pay… 36 arXiv — NLP / Computation & Language research 2mo ago AIGP: An LLM-Based Framework for Long-Term Value Alignment in E-Commerce Pricing arXiv:2606.26787v1 Announce Type: cross Abstract: Traditional dynamic pricing models in large-scale e-commerce suffer from limited interpretability, poor utilization of unstructured information, and misalignment with long-term business objectives such as cumulative Gross… 26 arXiv — Machine Learning research 2mo ago State Representation Matters in Deep Reinforcement Learning: Application to Energy Trading arXiv:2606.27032v1 Announce Type: new Abstract: Energy trading decisions depend not only on current market prices, but also on expected future market conditions, and operational constraints. This makes the state representation given to a reinforcement learning agent an important… 5 r/LocalLLaMA community 2mo ago Prices of graphic cards are going crazy, should I buy a second card though? A few months ago, I bought a RX 7900 XTX 24g to start toying with local LLM, at 900€ new. Little I knew that now I want to add a second card to my rig, but prices have gone insane! Adding a new 7900 XTX would cost me 1200€ as new now, used price is around 900€ now, and the last… 38 r/MachineLearning community 2mo ago [R] Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost Token-based billing is causing my company to reevaluate small language models. I came across this paper that shows SLM supervised fine-tuning on traces from orchestration of frontier models can be nearly as performant and much cheaper. Has any tried this in the real world?  … 34 r/LocalLLaMA community 2mo ago rtx 6000 pro owners, do you regret? I found the last dealership in my area that has rtx 6000 pro available, i already wanted to buy it 6 months ago when it was around $8k, now prices increased to $13k ish. Regardless the price, are you happy with it? I assume you are using qwen3.6 27b, is it worth it? Please share… 9 r/LocalLLaMA community 2mo ago New Apple Memory Prices https://preview.redd.it/00o5xtaznf9h1.png?width=696&format=png&auto=webp&s=60a3306ea86a9b0d1f58c435b7dbb0a42761a415 Apple raised the prices across the product line this morning:… 33 Hacker News — AI on Front Page community 2mo ago Apple raises prices of MacBooks, iPads https://9to5mac.com/2026/06/25/apple-price-increases-mac-ipa... Comments URL: https://news.ycombinator.com/item?id=48672732 Points: 362 # Comments: 565 7 arXiv — Machine Learning research 2mo ago MacroLens: A Multi-Task Benchmark for Contextual Financial Reasoning under Macroeconomic Scenarios arXiv:2606.24950v1 Announce Type: new Abstract: Financial decision-making is contextual: forecasting prices, valuing companies, and assessing event exposure weigh price history, accounting fundamentals, macroeconomic regime, and contemporaneous text. A benchmark over these four… 25 r/LocalLLaMA community 2mo ago Locked Dell quote for 6x RTX PRO 6000 Max-Q at $8,960 — expires tonight. What would you do? Building an inference cluster to run GLM 5.2 locally. Got a Dell quote locked at $8,959.99/unit for 6x RTX PRO 6000 Blackwell Max-Q (300W). List price just jumped to $15,999 yesterday. Quote expires in ~3 hours and I can't swing all 6 right now. I have a second quote for 2 units… 36 r/LocalLLaMA community 2mo ago MINISFORUM DEG1 Oculink eGPU Dock Refurbished - $59 I got one of these refurbished units last year. I have nothing but good things to say about it. It works great. It has heft to it to keep the GPU secure. And unlike some cheaper Oculink docks, it has redrivers for signal integrity.   submitted by   /u/fallingdowndizzyvr… 20 r/MachineLearning community 2mo ago I compiled LLM inference pricing across 7 providers — the caching numbers are surprising(spreadsheet included) [R] I've been comparing GPU/LLM providers for a side project and ended up with way too many browser tabs and spreadsheets. So I decided to pull the public pricing data into one sheet and compare it side by side. A quick disclaimer: this is not benchmark data . I didn't run latency… 32 arXiv — NLP / Computation & Language research 2mo ago Task Decomposition for Efficient Annotation arXiv:2606.24734v1 Announce Type: new Abstract: High-quality annotations of structured representations are expensive to collect over large corpora. Manual annotation of structure is laborious, and model-based annotation, although cheaper to generate, requires expensive… 24 r/LocalLLaMA community 2mo ago I want to add a second 7900XTX, question about pcie2/3/4 I've got a 7900XTX in my old gaming PC, now I want more vram and if I stay on one GPU I can only reasonably get 32GB and that just doesn't sound good enough. Using two slots, 48GB sounds way better and is much cheaper. I think 48GB is the minimum I want to have after going to… 26 r/LocalLLaMA community 2mo ago Openrouter model prices implying heavier quantization? Theres been a lot of talk about quiet quantization of models and what access to guaranteed model quality would look like. I’ve been trying to sanity check the economics of running large open models, and I’m having trouble making the numbers work. Take GLM-5.2 as an example. Even… 18 Page 5 of 9 · 410 articles ← Newer Older →