News / #hardware Tag Hardware 500 articles archived under #hardware · RSS Sign in to follow r/LocalLLaMA community 2mo ago Clustering 3x Jetson Nano Orin Supers Hey everyone! Recently, I released a blog on how to setup a cluster out of your Raspberry Pi 4bs and Mac minis for distributed training and inference Now its time to do the same with Jetson Nano Orin Super! Why ? - 1024 CUDA Cores (Ampere) - 8GB unified memory LPDDR5 - 6x ARM… 26 Hacker News — AI on Front Page community 2mo ago Google to pay SpaceX $920M a month for compute capacity at xAI data centers Article URL: https://www.cnbc.com/2026/06/05/google-to-pay-spacex-920-million-a-month-for-xai-compute-capacity.html Comments URL: https://news.ycombinator.com/item?id=48417490 Points: 231 # Comments: 804 38 Ars Technica — AI news-outlet 2mo ago "We pissed off a lot of people": Giant data center plan cut 50% amid protests Developer felt "beaten up," with "no choice" but to shrink data center. 35 r/MachineLearning community 2mo ago How do you identify researchers who are good? [D] About 10 years ago, I got into the basics of ML (like regression, KNN's, LVQ's) and read a few papers before taking a break a few years back. It feels like now, there's a lot of researchers in AI. How do you identify the ones who are actually solid vs those who (forgive my… 19 TechCrunch — AI news-outlet 2mo ago AirTrunk commits $30B to build 5GW of AI data centers in India The Australian data center operator plans to set up 5GW of capacity in India. 14 arXiv — Machine Learning research 2mo ago Staged Factorial Screening for Budget-Constrained Micro-Pretraining arXiv:2606.05186v1 Announce Type: new Abstract: Budget-constrained micro-pretraining often requires triaging many candidate recipes on a shared accelerator before larger search budgets are spent. We study whether a staged fractional-factorial workflow can recover stable early… 14 arXiv — Machine Learning research 2mo ago A Machine Learning-Based Framework for Discovering Huntington's Disease Stages: Integrating Graph Representation Learning and clustering to Uncover Progression Dynamics in Longitudinal Enroll-HD Dataset arXiv:2606.06196v1 Announce Type: new Abstract: Huntington's disease (HD) is a progressive brain disorder that gradually affects movement, cognitive function, and behavior. Identifying the stage of the disease accurately and consistently is important for understanding its… 31 The Information — AI news-outlet 2mo ago Data Center Developer Switch in Talks to Raise Billions at $50 Billion-Plus Valuation Data center developer Switch is in talks to raise billions of dollars at a valuation of at least $50 billion, a level that would make it one of the most valuable privately held data center operators, The Information reported late Thursday . Brookfield Asset Management, KKR and… 28 The Information — AI news-outlet 2mo ago Data Center Developer Switch in Talks to Raise Billions at $50 Billion-Plus Valuation Data center developer Switch is in talks to raise billions of dollars at a valuation of at least $50 billion, as it seeks to capitalize on soaring demand for the infrastructure needed to support artificial intelligence, according to people with knowledge of the deal. Brookfield… 34 r/MachineLearning community 2mo ago Scrap the LLMs. Scoring 4.76% on the brand new ARC-3 using pure code, a 2012 AMD CPU, and zero AI tokens.[P] Hey everyone, The ARC Prize 2026 just launched the interactive ARC-AGI-3 track, and the collective AI world is panic-renting massive H100 clusters trying to get multi-billion parameter LLMs to navigate these dynamic environments. Predictably, out-of-the-box LLMs are faceplanting… 31 TechCrunch — AI news-outlet 2mo ago Meta steals a tactic from Tesla and builds data centers in tents Meta may have one found one way to slash its massive data center bill: tents. 9 r/LocalLLaMA community 2mo ago Qwen 3.6 27B 30GB Same top p: 98.358 ± 0.033 % vs UD Q8 K XL 33GB Same top p: 97.426 ± 0.041 % This is not a diss to Unsloth, they make great quants and really move this community forward. I've been experimenting with quanting specific sublayers based on which ones have the most outliers post Q8 quant. I basically did a BF16 to Q8_0 conversion and looked at the post quant… 8 Ars Technica — AI news-outlet 2mo ago How some data center operators are tackling their water use problems Hyperscalers have come under scrutiny for their impact on water quality and availability. 7 The Information — AI news-outlet 2mo ago Fusion Startup Helion Nearly Triples Valuation to $15.5 Billion in Thrive-led Round Helion Energy, a nuclear fusion startup backed by OpenAI’s Sam Altman, still has to prove it can produce electricity to serve data centers and other customers. But investors seem confident it can deliver. The Everett, Wash.–based company said it has raised $465 million in… 33 arXiv — Machine Learning research 2mo ago Contrastive Learning and Correlation Clustering for Sequences of Network Telescope Data arXiv:2606.04733v1 Announce Type: new Abstract: Understanding activities of Internet scanners is challenging; it often requires identifying relationships between sources, a task for which semantic annotations are scarce. This work investigates whether semantically meaningful… 36 arXiv — NLP / Computation & Language research 2mo ago Arithmetic Pedagogy for Language Models arXiv:2606.05106v1 Announce Type: new Abstract: We investigate whether methods of human mathematics pedagogy can guide the training of language models toward arithmetic reasoning. Building on the GASING method -- an Indonesian pedagogy that solves basic arithmetic through a… 32 arXiv — Machine Learning research 2mo ago A Nonmonotone Gradient-Based Algorithm for Symmetric Nonnegative Matrix Factorization and Graph Clustering arXiv:2606.02887v1 Announce Type: new Abstract: Symmetric nonnegative matrix factorization (Symmetric NMF) approximates a matrix as $WW^T$ with nonnegative rectangular factor $W$. It has broad applications in graph clustering and machine learning. In contrast to the NMF,… 9 arXiv — Machine Learning research 2mo ago KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators arXiv:2606.02963v1 Announce Type: new Abstract: Production inference increasingly targets a heterogeneous mix of accelerators. Agentic pipelines interleave reasoning, tool calls, and multi-agent coordination, each with distinct compute and memory profiles. For optimal… 19 arXiv — NLP / Computation & Language research 2mo ago Clustered Self-Assessment: A Simple yet Effective Method for Uncertainty Quantification in Large Language Models arXiv:2606.03846v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate remarkable performance across diverse tasks, but they often generate responses that appear plausible while being factually incorrect. This problem is compounded by the lack of explicit… 13 r/LocalLLaMA community 2mo ago I Put a Datacenter GPU in My Gaming PC for £200 Hey there! I wrote a blogpost about my experience running local models on a V100 from a newbie perspective and got loads of views outside of reddit, so I thought I'd share it here too!   submitted by   /u/tymscar [link]   [comments] 33 The Information — AI news-outlet 3mo ago SK Hynix to Double Capacity as AI Strains Memory Supply SK Hynix plans to double its memory chip capacity within five years as AI demand keeps straining global supply, Bloomberg reported. The expansion can ease one of the biggest hardware constraints facing AI data centers. Chairman Chey Tae-won said in Taipei that the memory crunch… 20 arXiv — Machine Learning research 3mo ago PE-means: Improved Differentially Private $k$-means Clustering through Private Evolution arXiv:2606.00342v1 Announce Type: new Abstract: We study the problem of differentially private (DP) $k$-means clustering in Euclidean space. Previous solutions rely on summing the private data directly, which induces a sensitivity proportional to the domain. We introduce… 17 arXiv — NLP / Computation & Language research 3mo ago French parsing enhanced with a word clustering method based on a syntactic lexicon arXiv:2606.00634v1 Announce Type: new Abstract: This article evaluates the integration of data extracted from a French syntactic lexicon, the Lexicon-Grammar (Gross, 1994), into a probabilistic parser. We show that by applying clustering methods on verbs of the French Treebank… 16 arXiv — NLP / Computation & Language research 3mo ago Agentic Clustering: Controllable Text Taxonomies via Multi-Agent Refinement arXiv:2606.01255v1 Announce Type: new Abstract: Recent text-clustering methods use large language models to propose a cluster taxonomy from a corpus and then assign each text to it. These pipelines are fundamentally programmatic: the sequence of LLM calls and the rules for… 37 NVIDIA Developer Blog official-blog 3mo ago Run Local AI Agents with Faster Models and Multi-Node Clustering on NVIDIA DGX Spark The rise of autonomous, long-running AI agents has introduced a new class of compute demand, namely tasks that maintain large context windows, spawn concurrent... 16 TechCrunch — AI news-outlet 3mo ago Water access is now a risk factor in SpaceX’s IPO The company says it needs "significant" water resources to cool its data centers, and that access to abundant, affordable water is a challenge. 14 Hacker News — AI on Front Page community 3mo ago AI Agent Guidelines for CS336 at Stanford Article URL: https://github.com/stanford-cs336/assignment1-basics/blob/main/CLAUDE.md Comments URL: https://news.ycombinator.com/item?id=48359232 Points: 207 # Comments: 92 8 The Information — AI news-outlet 3mo ago SpaceX Discloses Updated Details of Anthropic Deal, Flags Data Center Water Risks SpaceX’s compute rental deal with Anthropic will run for a minimum of six months, according to an amended securities filing on Monday, providing an additional level detail that had not been included in the IPO filing the company made public in May. Anthropic’s agreement to pay… 19 OpenAI official-blog 3mo ago Building the infrastructure for the Intelligence Age in Michigan OpenAI breaks ground on a 1GW data center project in Michigan as part of Stargate, building AI infrastructure to expand access, create jobs, and support communities. 17 arXiv — Machine Learning research 3mo ago ScaleMAP: Preserving Local Density and Neighborhood Structure in Low-Dimensional Embeddings arXiv:2605.30597v1 Announce Type: new Abstract: Nonlinear dimensionality-reduction methods such as UMAP and PaCMAP adaptively normalize local distances during graph construction, erasing neighborhood scale from the data. This distorts more than relative cluster sizes: sparse… 9 arXiv — NLP / Computation & Language research 3mo ago Pairwise Reference Alignment as a Model-Level Ordinal Observable arXiv:2605.30758v1 Announce Type: new Abstract: Pairwise preference data is widely used in language-model evaluation and alignment, often for model ranking, reward modeling, or preference optimization. This note formulates a more basic measurement question: given a reference… 18 OpenAI official-blog 3mo ago “Data Center Bandwagon” Campaign: US-targeted influence activity OpenAI banned a likely PRC-origin cluster using AI to generate social media content criticizing US data centers and AI infrastructure. 33 r/LocalLLaMA community 3mo ago G7 agrees on shared language around open-source AI and open weights AI Basically stuff we already knew here, but now governments understand it too. I found the news here: https://www.phoronix.com/news/G7-On-Open-Source-AI   submitted by   /u/Kahvana [link]   [comments] 16 TechCrunch — AI news-outlet 3mo ago Erin Brockovich takes aim at data center secrecy Environmental activist Erin Brockovich has a new mission. 35 r/LocalLLaMA community 3mo ago Added an old 2070 Super to my rig and I can't go back...worse, now I need more Context: I built a new system last year November before everything went to shit. I spent like 5k for a 5090, 9800X3D and 96GB RAM. Recently (last 2-3 months) I'm heavily working on my local setup. Ditched Windows, went Ubuntu > Manjaro > CachyOS (now) and I'm basically building… 36 Hacker News — AI on Front Page community 3mo ago I put a datacenter GPU in my gaming PC Article URL: https://blog.tymscar.com/posts/v100localllm/ Comments URL: https://news.ycombinator.com/item?id=48345694 Points: 241 # Comments: 154 5 r/MachineLearning community 3mo ago Built an AI Accelerator and opensourced it. [P] There is a huge gap in open source AI accelerators, so I implemented mine . Popular and well known ones are already legacy and doesn't support contemporary operations like Attention. Here is what makes mine special: Attention mechanism smelted directly into silicon Prototyped… 25 r/LocalLLaMA community 3mo ago DIY Local 2x DGX Spark cluster cooler with automatic temperature controlled fan. I’ve found that DGX Sparks can get pretty warm when you cluster them together. You are forced to keep them close together because the ConnectX-7 cable made for these is extremely short )like less than a foot). I have both a DGX Spark Founder’s Edition and a GIGABYTE AI TOP Atom… 32 r/MachineLearning community 3mo ago How would you model this "strand" clustering problem? [P] https://preview.redd.it/llqlupnwng4h1.png?width=2188&format=png&auto=webp&s=7fae5860babaffa1c8bfdcb1468b374eb38ac55d I'm currently building a computer vision application. I've managed to successfully train a YOLO model to detect the object I'm interested in for my videos. The… 33 r/LocalLLaMA community 3mo ago Dell confirms XPS laptop with NVIDIA N1X at Computex ( basically a DGX Spark GB10 for consumers with Windows )   submitted by   /u/fallingdowndizzyvr [link]   [comments] 24 r/LocalLLaMA community 3mo ago My home data center System 1: Threadripper 3960x 24c 4x 3090 ti 128gb ddr4 System 2: Xeon 8352 36c 4x 5070 ti 128gb ddr4 System 3: Intel 14700k 24c 64gb ddr5 5090 System 4: Ryzen 5950x 16c 64gb ddr4 2x 5070 ti The first system uses two PSUs to handle the almost 2000w full load of the 3090s. Was… 4 TechCrunch — AI news-outlet 3mo ago SoftBank says it will invest up to €75 billion to build French data centers The goal, the firm said, is to develop and operate up to 5 gigawatts of additional data center capacity. 30 The Information — AI news-outlet 3mo ago Softbank to Invest Up To 75 Billion Euros on AI Data Centers in France SoftBank Group announced a commitment to develop and operate five gigawatts of AI data center capacity in France, with an investment of up to 75 billion euros, or about $87.5 billion. The commitment is SoftBank’s largest AI infrastructure investment to date in Europe, the… 13 r/LocalLLaMA community 3mo ago Why does Thinking Output More Tokens Than a Response? I was too lazy to use a vector DB + Embedding + Clustering for this list of 1000 items I wanted to categorize. I was hoping to use a local LLM to do it, but it would only respond with a list of about 100 items or so and their categories. It confused me because when I saw the… 22 r/LocalLLaMA community 3mo ago made a local voice AI for windows you can talk to in any language. open source, bring your own key been building this on and off for a while and finally got it to a point where i'm not embarrassed to share it, so here goes. it's called Shadow AI. basically a voice-first AI companion that runs on your own windows machine. you just talk to it and it talks back, no typing… 38 r/LocalLLaMA community 3mo ago Anyone using Flash Attention 2 (ai-bond) on their V100's? How is the performance? I just Installed Flash Attention 2 from here: https://github.com/ai-bond/flash-attention-v100 " I did some basic benchmarks and I am getting from 4x-7x memory utilization. However, benchmarks don't always translate to real world scenarios. **I have noticed that the thinking time… 30 r/LocalLLaMA community 3mo ago We gave a Reachy Mini a real-time voice brain We attended an event the other day and found this little guy lying on our desk, a Reachy Mini from Hugging Face. It belongs to the daughter of the event organizer. We got curious about how it worked, and an hour later we'd given it a brain. The model basically becomes Reachy. It… 19 llama.cpp releases dev-tools 3mo ago b9402 hexagon: basic/generic op fusion support and RMS_NORM+MUL fusion ( #23835 ) Updating infra to enable op fusion and using RMS_NORM+MUL as the use-case. macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework… 17 arXiv — Machine Learning research 3mo ago Cycle-Space Informed Detection of Autoencoded Blind False Data Injection Attacks on Power Systems arXiv:2605.28912v1 Announce Type: new Abstract: The rapid growth of AI-driven data centers and large-scale energy storage systems is increasing the reliance of power system operation on real-time measurement data and automated decision-making. However, many existing detection… 28 arXiv — Machine Learning research 3mo ago Cluster-Level Attention-Guided Parallel Decoding for Masked Diffusion Language Models arXiv:2605.29607v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) enable parallel decoding by predicting all masked positions at each denoising step, yet existing training-free samplers usually decide which positions to commit at token-level granularity.… 28 Page 9 of 10 · 500 articles ← Newer Older →