News / #hardware Tag Hardware 500 articles archived under #hardware · RSS Sign in to follow arXiv — Machine Learning research 1mo ago Data-Native Global Optimization for Big Data K-means Clustering arXiv:2607.15835v1 Announce Type: new Abstract: Big data clustering remains challenging: the Minimum Sum-of-Squares Clustering (MSSC) problem underlying K-means is NP-hard, and existing methods either reach poor local minima or require prohibitive metaheuristic hybrids. We… 32 arXiv — Machine Learning research 1mo ago Understanding Reasoning from Pretraining to Post-Training arXiv:2607.16097v1 Announce Type: new Abstract: Reinforcement learning (RL) has become central to improving large language models (LLMs) on complex reasoning tasks, yet RL post-training is largely studied in isolation from the pretraining that precedes it. As a result, two basic… 20 arXiv — NLP / Computation & Language research 1mo ago An MLIR-Based Compilation Method for Large Language Models arXiv:2607.15865v1 Announce Type: new Abstract: Large Language Models (LLMs) have become the dominant workload on modern AI accelerators, yet deploying them on specialized hardware still faces two core challenges: how to import a trained model into a compiler-friendly… 20 Hugging Face Daily Papers research 1mo ago Understanding Reasoning from Pretraining to Post-Training Abstract Reinforcement learning (RL) has become central to improving large language models (LLMs) on complex reasoning tasks, yet RL post-training is largely studied in isolation from the pretraining that precedes it. As a result, two basic questions remain open: (1) how do… 38 r/LocalLLaMA community 1mo ago Introducing ASCIITermDraw Bench | Testing the ability of VLMs to Generate and Edit ASCII ASCIITermDraw-Bench: Can a Model Actually Draw in ASCII? Do we really need a image generator to relay our thoughts about - an architecture? a topology? a cluster og N nodes? Is it possible to let our AI assistants, easily absorb and understand and make possible changes easily… 8 r/LocalLLaMA community 1mo ago The AMD Instinct MI350P is a HBM PCIe AI Accelerator That Has Been All Over   submitted by   /u/Neurrone [link]   [comments] 33 r/LocalLLaMA community 1mo ago Typical_p + Qwen There's a lot of people using Qwen here. I finally remembered what I wanted to do: typical_p has been merged eons ago and it's a sampler that was basically designed for countering repetition loops. I didn't yet try it, but I think that should help with its looping hiccups… 12 r/LocalLLaMA community 1mo ago When will we get more small LLMs? Basically the title. We had our last drop in the beginning of April, do we just not get a refresh of Gemma or Qwen?   submitted by   /u/Aggravating-Push-207 [link]   [comments] 12 r/LocalLLaMA community 1mo ago Welcome to 2014 - my new rig Hi all, Probably for some others me, there is a more conservative budget when it comes to their AI hobby. I've been collecting basically e-waste and could now assemble something partially working from them. Specs: $ neofetch .-/+oossssoo+\-. molbal@... ´:+ssssssssssssssssss+:`… 34 TechCrunch — AI news-outlet 1mo ago Why the first GPU financiers are turning to inference chips in a $400 million deal A $400 million chip-backed loan points to the next wave of AI infrastructure deals. 27 arXiv — Machine Learning research 1mo ago Counterfactuals for Feature-Weighted Clustering arXiv:2607.14719v1 Announce Type: new Abstract: Counterfactual explanations provide local, interpretable insight by identifying changes to an input that would alter its assigned outcome. Although well established in supervised learning, their extension to clustering is less… 12 arXiv — Machine Learning research 1mo ago Analysis of Public Schools Educational Performance Based on Causal Models and Hierarchical Clustering arXiv:2607.14124v1 Announce Type: cross Abstract: The increasing availability of large-scale educational datasets has expanded the use of quantitative methods for investigating school performance. However, institutional heterogeneity among schools and the structural complexity… 23 Simon Willison community 1mo ago Spot birds not golf Suggestion for hyperscalers feeling pressure over data center water use: Buy up a few exclusive country clubs, convert the golf courses into public parks, pay for guides and binoculars to get the previous members into birdwatching - help them embrace a more sustainable hobby!… 27 r/LocalLLaMA community 1mo ago Added SearXNG and I don't even know what to say anymore. I just have to show someone other than the guys at work who think I'm crazy. I've been working on this app since about early June of 2025 and while it has come a long way in its workspace tools, I had always only had basic web capability. But I saw some posts yet again recently… 11 TechCrunch — AI news-outlet 1mo ago Roblox launches an AI-powered game creation feature in its mobile app Roblox's new "Build" feature lets users generate basic games using a single text prompt. 37 r/LocalLLaMA community 1mo ago Kimi K3 Release Video [Made with Kimi K3] Everyone's probably seen the remotion thing that went viral a couple months back with CC. Its basically that with Kimi K3 as the model provider. Prev. example with GLM 5.2: https://www.reddit.com/r/LocalLLaMA/comments/1u8kyqf/glm_52_release_video_made_with_glm_52/ Feels much… 33 Latent.Space news-outlet 1mo ago 🔬 The Lab of the Future Should Feel Like a Data Center — Andy Beam & Rafa Gómez-Bombarelli, Lila Sciences Lila is betting that science, not the internet, is the last untapped source of training data. We went to find out what that actually looks like in a room full of robots. 30 arXiv — Machine Learning research 1mo ago EMAGN: Efficient Multi-Attention Graph Network via Learned Clustering for Scalable Traffic Forecasting arXiv:2607.13241v1 Announce Type: new Abstract: Traffic forecasting is highly challenging due to complex and nonlinear spatial and temporal dependencies. Self-attention mechanisms have been widely adopted to model dynamic and long-range dependencies, achieving state-of-the-art… 17 arXiv — Machine Learning research 1mo ago Agora: Collective and Permissionless Internet-Scale Pretraining of Large Language Models arXiv:2607.13332v1 Announce Type: new Abstract: Training large language models at the multi-billion to trillion parameter scale is confined to datacenters, where data-parallel (DP) and model-parallel (MP) techniques presume homogeneous accelerators, high-speed interconnects, and… 24 arXiv — Machine Learning research 1mo ago Clustering algorithms for multivariate wind farm SCADA data filtering arXiv:2607.13544v1 Announce Type: new Abstract: During wind farm operation, Supervisory Control and Data Acquisition (SCADA) systems record numerous anomalies, transients, and specific operational modes, leading to large datasets. However, for a wide range of applications, only… 31 arXiv — Machine Learning research 1mo ago Optimal and Efficient Contextual Combinatorial Semi-bandits with General Function Approximation arXiv:2607.13686v1 Announce Type: new Abstract: We study the contextual combinatorial semi-bandit (CCSB) problem with general reward function approximation. At each round, the learner observes a context, selects a combinatorial action consisting of a subset of basic arms, and… 23 arXiv — Machine Learning research 1mo ago Hierarchical $\mathcal{F}$-Clustering: Approximation and Hardness of Clustering into Trees and Bounded Diameter Graphs arXiv:2607.13217v1 Announce Type: cross Abstract: Consider the following variation on the Hierarchical Clustering problem: Usually, while building a hierarchical clustering, one recursively partitions the data until each cluster becomes a singleton. We relax the halting… 4 r/LocalLLaMA community 1mo ago RL post-training on 14 Macs across 4 countries Disclosure: I work at Pluralis Research, the lab that built this. Code is open, and I'm happy to answer questions. TL;DR: As far as we can tell, this is the first RL post-training run whose entire rollout fleet ran on consumer Macs over the open internet. Setup 14 Macs across 4… 19 Hacker News — AI on Front Page community 1mo ago Mysteries of Telegram Data Centers (2022) Article URL: https://dev.moe/en/3025 Comments URL: https://news.ycombinator.com/item?id=48920475 Points: 221 # Comments: 110 7 arXiv — Machine Learning research 1mo ago Cluster-Weighted EDMD arXiv:2607.12243v1 Announce Type: new Abstract: Extended Dynamic Mode Decomposition (EDMD) approximates Koopman operators from data, but a single global operator is inefficient when different state-space regions exhibit distinct local dynamics. We introduce Cluster-Weighted EDMD… 13 arXiv — Machine Learning research 1mo ago Energy-Based Physics-Informed Form Finding for Clustered Tensegrity Structures arXiv:2607.12888v1 Announce Type: new Abstract: Tensegrity form-finding and physical property prediction are fundamental inverse problems in structural mechanics, which aim to determine equilibrium configurations and internal force distributions. These problems are challenging… 27 Hacker News — AI on Front Page community 1mo ago Jurassic Park computers in excruciating detail Article URL: https://fabiensanglard.net/jurrasic_park_computers/index.html Comments URL: https://news.ycombinator.com/item?id=48915709 Points: 277 # Comments: 62 7 r/MachineLearning community 1mo ago Things I got wrong building an incremental indexing pipeline [P] I've been working on incremental indexing pipelines lately, basically keeping a vector store in sync as the source data changes, and I keep finding the same bugs never show up until it's been running a while. Biggest one for me is deletes. I tested the "new doc comes in, gets… 10 TechCrunch — AI news-outlet 1mo ago OpenAI’s new flagship model deletes files on its own, people keep warning A number of social media posts claim that GPT-5.6 Sol deleted files and data without warning. OpenAI had basically disclosed the problem in June. 11 Hugging Face Daily Papers research 1mo ago A Theory of Contrastive Learning with Natural Images Abstract Why does contrastive learning with simple images and augmentations yield useful representations for downstream tasks? We address this question by analytically computing the optimal representation in terms of a contrastive loss for a range of basic augmentations and any… 33 TechCrunch — AI news-outlet 1mo ago New York State halts construction of all new data centers New York has become the first state to temporarily halt approval of large data centers, as Gov. Kathy Hochul argues the AI-driven building boom shouldn’t come at the expense of higher electricity costs, water supplies, or local control. 27 Ars Technica — AI news-outlet 1mo ago New York bans data center construction for a year, rattling AI industry New York’s data center moratorium may become the blueprint for anti-AI movement. 17 TechCrunch — AI news-outlet 1mo ago Sam Altman’s space data center trash talk is what most experts already believe "homeboy you're the one sellling [sic] public market investors on short-term space datacenters." 5 arXiv — Machine Learning research 1mo ago Group Invariant Spectral Embedding arXiv:2607.08987v1 Announce Type: new Abstract: Spectral embedding methods are widely used for dimensionality reduction and clustering of high-dimensional datasets with intrinsic low-dimensional structures. Although many datasets of practical interest exhibit invariance under… 19 arXiv — NLP / Computation & Language research 1mo ago Relation Extraction Model Based on Semantic Enhancement Mechanism arXiv:2311.02564v2 Announce Type: replace Abstract: Relational extraction is one of the basic tasks related to information extraction in the field of natural language processing, and is an important link and core task in the fields of information extraction, natural language… 9 r/LocalLLaMA community 1mo ago Has anyone gotten Llama.cpp (or other) working using Intel iGPU (arrowlake) where it actually improves anything? I Recently did a bunch of tests and wrote them all up on here, but the short version is that Vulkan basically doesn't work (or when it does, it's at 1tok/s at best). SYCL works pretty well, seems to run the Qwen3.6 35b models at around 12tok/s. The prefill part is around 20tok/s… 17 r/LocalLLaMA community 1mo ago If you use Open Code or other agenting programs you are leaving a lot of t/s if you don't actually use agents in parallel. Benchmark : RTX5090, Qwen3.6 35B loaded via LM studio with parallel tasks set to 8 As many of you know t/s is super important. It's how fast your stuff gets done. I create via open code benchtest and run it. Thanks to it i know that if i don't run at least 4 agents i basically leave HALF of performance. So whatever you do single project in open code that uses… 34 llama.cpp releases dev-tools 1mo ago b9957 server: improve tools, remove apply_diff ( #25498 ) server: improve tools, remove apply_diff improve edit tool add tools_io abstraction add tools_io_basic fix build move utils to class member add const macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI… 9 llama.cpp releases dev-tools 1mo ago b9949 opencl: cluster-parallel decode FA for Adreno ( #25473 ) macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu arm64 (CPU) Ubuntu s390x (CPU) Ubuntu x64 (Vulkan) Ubuntu arm64… 37 arXiv — Machine Learning research 1mo ago Structure Learning on Clustered Data arXiv:2607.08238v1 Announce Type: new Abstract: Recent algorithmic advances have made directed acyclic graph (DAG) structure learning scalable for causal discovery. Yet, the currently available techniques assume a completely homogeneous population, precluding their application… 6 arXiv — Machine Learning research 1mo ago CASL-VAE: Learning Structured Latent Variables from Unpaired Data for Semi-supervised Clustering and Paired Sample Generation arXiv:2607.08254v1 Announce Type: new Abstract: Quantifying variability in a target population relative to a reference population is central to many scientific and clinical problems (e.g., diseased vs. healthy). Yet, without paired data and in the presence of heterogeneous… 12 arXiv — NLP / Computation & Language research 1mo ago A Multi-cluster Boundary Learning Method for Out-of-Scope Intent Detection via MiniLM Embedding arXiv:2607.07974v1 Announce Type: new Abstract: Intent detection is a critical task that bridges human intents and system actions in human-machine interaction systems. However, there still exist challenges for detecting out-of-scope (OOS) intents. (i) The traditional methods… 35 r/LocalLLaMA community 1mo ago Qwen 3.6 Q2-FP8 Terminal Bench 2 and GPQA Scores TL;DR: Quantization has a marked impact on agentic performance but little effect on knowledge. I manage a small HPC cluster at a university, and we have recently begun running common benchmarks to help our users understand the effects of quantization. We have just completed the… 34 r/LocalLLaMA community 1mo ago Exploring FlashAttention-3/4 optimizations on RTX GPUs I was curious whether any of the FA-3/4 optimizations transfer to RTX GPUs. vLLM/SGLang attention falls back to FA-2 on consumer cards (FA-3 and FA-4 are datacenter-only), so I wanted to know if there's any performance left on the table, and I rebuilt the attention kernels from… 21 Hacker News — AI on Front Page community 1mo ago No leap second will be introduced at the end of December 2026 Article URL: https://datacenter.iers.org/data/latestVersion/bulletinC.txt Comments URL: https://news.ycombinator.com/item?id=48846281 Points: 209 # Comments: 167 18 arXiv — Machine Learning research 1mo ago Converge to Surprise: Evolutionary Self-supervised Image Clustering arXiv:2607.06887v1 Announce Type: new Abstract: Most self-supervised image clustering models, actually almost all deep learning approaches, are based on gradient descent: In order to calculate the loss, every optimization step requires a clearly defined target, whether a… 17 arXiv — Machine Learning research 1mo ago Imputation Meets Clustering: Exploiting Latent Subgroup Structure for Missing Data Recovery arXiv:2607.06930v1 Announce Type: new Abstract: Missing data is prevalent in practical applications, making effective imputation an essential preprocessing step for downstream analysis. Real-world datasets often exhibit complex latent structures composed of multiple subgroups… 18 arXiv — Machine Learning research 1mo ago FMMVCC: Fuzzy Mamba-based Multi-View Contrastive Clustering for Univariate Time Series arXiv:2607.07258v1 Announce Type: new Abstract: In many realistic scenarios, large volumes of time series data are generated with limited or expensive annotations. This limitation makes supervised learning methods difficult to apply and leads to the use of unsupervised… 25 r/LocalLLaMA community 1mo ago What GUI-first coding tool tool are you pairing your local LLMs with? Opencode isn't it for me. I've grown very frustrated with OpenCode. The web GUI and desktop app ideas are good, but the execution not so much. The GUI is lacking so many basic features. It's clear that the TUI is more important to the devs. Is there anything free that provides a more feature-rich GUI?… 14 r/MachineLearning community 1mo ago Why does the same H100 cost 5x more depending on where you rent it? [D] I kept finding wildly different prices for the same GPU across providers and data centers, so I built a OS CLI that searches live GPU capacity and shows the cheapest available routes npx gpu-price-finder Supports RTX 4090, RTX 5090, L40S, A100, H100 and lets you filter by… 20 Page 5 of 10 · 500 articles ← Newer Older →