News / #outage Tag Outages 158 articles archived under #outage · RSS Sign in to follow Marcus on AI community 1d ago 5 lessons from the OpenAI / Hugging Face incident Did OpenAI really do the best they could? 8 arXiv — Machine Learning research 2d ago Beyond Capability Benchmarks: Learning Operational Fingerprints of LLM Cloud Services from Production Incident Metadata arXiv:2608.26332v1 Announce Type: new Abstract: Managed LLM services are now part of real production systems, but model selection and service planning still rely heavily on capability benchmarks that reveal little about operational behavior after deployment. We present… 26 arXiv — NLP / Computation & Language research 2d ago When Is Noise Response Universal? Tokenization as the Hidden Variable in Language Models arXiv:2608.26319v1 Announce Type: new Abstract: The performance of textual neural models often degrades when their inputs are corrupted by noise such as typos, OCR errors, or dropped words. We study the degradation rate across neural models, both sentence embeddings and… 18 TechCrunch — AI news-outlet 2d ago Here’s all the times AI has gone rogue and hacked other companies A recap of all the incidents involving LLMs made by Anthropic, Meta, and OpenAI, which went rogue and attacked real companies and individuals on the internet. 32 arXiv — Machine Learning research 3d ago FedQoS: Federated QoS-Risk Learning for Heterogeneous Indoor-Outdoor Access Selection arXiv:2608.25496v1 Announce Type: new Abstract: Reliable access selection in dynamic and heterogeneous indoor-outdoor environments is challenging because instantaneous radio measurements alone cannot capture future QoS degradation caused by mobility, blockage, traffic load, and… 33 Latent.Space news-outlet 3d ago [AINews] NVIDIA buys HuggingFace for $13B, as OpenAI publishes their HF incident retro Open Source wins! 10 Hacker News — AI on Front Page community 3d ago GitHub Outage Tracker: Is GitHub Cooked? Article URL: https://isgithubcooked.com/ Comments URL: https://news.ycombinator.com/item?id=49454728 Points: 206 # Comments: 131 9 Hacker News — AI on Front Page community 3d ago The Hugging Face incident and the road ahead Article URL: https://openai.com/index/hugging-face-incident-and-the-road-ahead/ Comments URL: https://news.ycombinator.com/item?id=49454314 Points: 214 # Comments: 258 38 TechCrunch — AI news-outlet 3d ago OpenAI releases its official report on the Hugging Face breach The report, which spans several discrete cybersecurity compromises, is the most complete accounting of the incident to date. 7 Hacker News — AI on Front Page community 3d ago Disruption with Some GitHub Services Article URL: https://www.githubstatus.com/incidents/hcbtzksccj2f Comments URL: https://news.ycombinator.com/item?id=49450722 Points: 224 # Comments: 134 36 arXiv — Machine Learning research 4d ago Data Leakage Inflates Generalizability of Power Outage Prediction Models arXiv:2608.24665v1 Announce Type: new Abstract: Power outage prediction models are increasingly used in assessments of climate-driven infrastructure risk, yet current evaluation practices obscure whether these models generalize to the novel conditions such applications require.… 6 OpenAI official-blog 4d ago The Hugging Face incident and the road ahead OpenAI shares findings from the Hugging Face security incident and the steps we’re taking to strengthen AI model security, monitoring, and alignment. 21 arXiv — NLP / Computation & Language research 5d ago Evidence-State Reliability Under Controlled Degradation: Parser-Validity Divergence in a Multi-Stage LLM Pipeline arXiv:2608.21559v1 Announce Type: new Abstract: Multi-stage LLM pipelines can remain structurally valid even when evidence available to downstream stages becomes incomplete, compressed, or conflicting. This paper introduces and operationalizes Evidence-State Reliability (ESR),… 6 arXiv — NLP / Computation & Language research 5d ago Improving Few-Step Language Flows with Untied Self-Conditioning arXiv:2608.22244v1 Announce Type: new Abstract: Flow-matching language models refine all token positions in parallel and can trade sampling steps for latency, yet generation quality still degrades sharply with few sampling steps. We trace a source of this degradation to a… 12 Hacker News — AI on Front Page community 9d ago The August 17 outage Article URL: https://github.blog/news-insights/company-news/the-august-17-outage-and-the-work-ahead/ Comments URL: https://news.ycombinator.com/item?id=49378957 Points: 499 # Comments: 568 15 r/LocalLLaMA community 9d ago Getting better at coding doesn't make a model better at everything else A majority of users in this sub use LLMs for coding/agentic tasks and I see why a lot of value is put into them but many try to say "Well coding has improved therefore it can just use tool calling and/or just look up what the user needs if there's a degradation for general… 37 r/LocalLLaMA community 10d ago Qwen3.8-23B-Mini-Me: A Depth-Pruned Qwen3.8-27B (to ~22.7BB) I've been working on a depth pruning approach and decided to try it out on the new Qwen3.8-27B model. I managed to get the model down to about 22.7B params without severe reasoning degradation. No fine-tuning was done, just strategic removal of layers. It's been working well for… 13 arXiv — NLP / Computation & Language research 11d ago AISA: AI Safety Assistant Framework for Continuous Improvement of Highway Construction arXiv:2608.17184v1 Announce Type: new Abstract: Job Safety Analysis (JSA) and pre-task planning can benefit from prior incident records, yet historical accident data is often stored as unstructured narratives that are difficult to consult at the point of planning. A novel… 15 Vercel — AI dev-tools 11d ago $1 million hacker challenge for Vercel Sandbox Agents need to run untrusted code, and the microVM has become the standard way to do it: a dedicated guest kernel per workload, isolated from the host and from every other workload on the same machine. But recent security research and real-world incidents have revealed that… 26 arXiv — Machine Learning research 12d ago Early Cycle Charge Trajectory Generative Prediction and Full Life Cycle Health Management of Iron-Chromium Flow Batteries Based on FlowBD-E1 arXiv:2608.14637v1 Announce Type: new Abstract: Long-duration stationary energy storage requires batteries whose degradation can be detected before substantial capacity loss has accumulated. Iron-chromium redox flow batteries are attractive for this role because they use… 29 arXiv — Machine Learning research 12d ago Real-Time State-of-Health Estimation and Online Degradation Prognosis from Partial Battery Discharge Using Physics-Informed Neural Networks arXiv:2608.14764v1 Announce Type: new Abstract: With the increasing integration of renewable energy sources, energy storage systems have become essential, making the accurate estimation of their State of Health (SOH) and degradation behavior critical. In this work, we propose a… 9 arXiv — NLP / Computation & Language research 12d ago Why Vision Fails as a Universal Bridge: Rectifying Modality Asynchrony in Multilingual MLLMs arXiv:2608.15085v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) exhibit substantial performance degradation in non-English visual reasoning, despite the strong multilingual competence of their text-only backbones. While mechanistic evidence from… 38 Hugging Face Daily Papers research 12d ago Ventor-QTest: Threat-Model-Driven Verification of Vendor-Hosted LLM APIs Abstract Ventor-QTest audits hosted open-weight model APIs via repeated and long-sequence black-box probes, measuring average and extreme fidelity loss to detect degradation in long-horizon agentic performance. Generated by thinkingmachines/Inkling-Small As large language models… 35 Hacker News — AI on Front Page community 12d ago Incident with Github.com Article URL: https://www.githubstatus.com/incidents/zkxwbgr0cnmx Comments URL: https://news.ycombinator.com/item?id=49330684 Points: 427 # Comments: 347 10 arXiv — Machine Learning research 17d ago GCPO: Diagnosing and Constraining Subspace Geometry in Rollout RL for LLMs arXiv:2608.11674v1 Announce Type: new Abstract: On-policy rollout methods such as GRPO are central to post-training of large language models, yet they frequently suffer from training instabilities, cross-task capability degradation, and response-length inflation. Although prior… 34 arXiv — NLP / Computation & Language research 17d ago Language-Conditional Dequantization: Recovering What Quantization Steals from Non-English Languages arXiv:2608.11786v1 Announce Type: new Abstract: Aggressive quantization disproportionately harms multilingual capability: in the sub-4B INT3 GPTQ regime, we measure 2-4x larger perplexity degradation on non-English languages than on English. We propose Language-Conditional… 28 arXiv — Machine Learning research 18d ago SQuaT: Self-Supervised Knowledge Distillation via Student-Aware Quantized Teacher Features arXiv:2608.10709v1 Announce Type: new Abstract: Quantization-Aware Training (QAT) enables the deployment of quantized models with minimal accuracy degradation. However, in practical scenarios, training labels are often unavailable due to privacy, copyright, or cost constraints.… 5 arXiv — NLP / Computation & Language research 18d ago The Multilingual Quantization Tax: Structural Collapse and Typological Fragility in Edge SLMs arXiv:2608.09941v1 Announce Type: new Abstract: While 4-bit weight quantization is critical for deploying Small Language Models (SLMs) on edge devices, evaluations of the resulting performance degradation-the quantization tax-remain overwhelmingly English-centric. We present a… 30 arXiv — Machine Learning research 20d ago MiCoPro: End-to-End Mixed Precision HW/SW Co-design with HW-aware Proxy Model arXiv:2608.06916v1 Announce Type: new Abstract: Quantized Neural Networks~(QNN) with low-bitwidth data have proven promising in efficient storage and computation on edge devices. To mitigate accuracy degradation while maximizing speedup, layer-wise mixed-precision… 31 Simon Willison community 22d ago Now we have a timeline of the OpenAI accidental attack against Hugging Face OpenAI gave a last-minute presentation at the Black Hat security on Wednesday about "the Hugging Face Incident" ( previously on this blog). The video was published yesterday. It's short and information dense and well worth watching, in particular because it provides full details… 28 r/LocalLLaMA community 23d ago Black Hat USA 2026: The 'Breaking' News: The OpenAI–Hugging Face Incident   submitted by   /u/SilentLennie [link]   [comments] 38 Smol AI News news-outlet 23d ago not much happened today **OpenAI** escalates its upcoming **Astra** model to "critical" cyber status due to significant advancements in agentic coding and cybersecurity, pausing some activities to strengthen controls. The "Hugging Face incident" highlights persistent multi-agent coordination failures… 8 arXiv — Machine Learning research 23d ago Equipment-centric workpiece localization in near real-time using deep learning-based vision and event-driven finite state machines arXiv:2608.05744v1 Announce Type: new Abstract: Continuous workpiece localization is essential for traceability and process coordination in hot forging, but direct tracking is unreliable because of extreme temperatures, surface degradation, and irregular routing. This study… 13 arXiv — Machine Learning research 23d ago Failing Gracefully: Mitigating Impact of Inevitable Robot Failures arXiv:2608.05313v1 Announce Type: cross Abstract: Service robots operate in household environments shared with humans, pets, and everyday objects, where they are highly susceptible to failures such as software crashes, hardware degradation, or unpredictable interactions. While… 25 Hugging Face Daily Papers research 23d ago ChronoVision: Temporal Reasoning via Latent State Reconstruction Abstract Multimodal large language models excel at passive perception but struggle with complex visual cognitive tasks requiring multi-step temporal reasoning. This degradation largely stems from the inherent ambiguity of language-based reasoning, which often fails to accurately… 15 Hacker News — AI on Front Page community 23d ago GitHub Actions and Pages are experiencing degraded availability Article URL: https://www.githubstatus.com/incidents/qcvjkzcs7j74 Comments URL: https://news.ycombinator.com/item?id=49198302 Points: 202 # Comments: 179 19 arXiv — NLP / Computation & Language research 24d ago The Fairness Collapse Phenomenon: Bias Amplification in Language Models Trained on Synthetic Data arXiv:2608.04268v1 Announce Type: new Abstract: Generative models trained on artificially generated data have been shown to exhibit model collapse, resulting in significant performance degradation. As synthetic content increasingly contaminates the training corpora of language… 8 arXiv — NLP / Computation & Language research 24d ago MIDAS: Multi-LLM Iterative Data-Adaptive Summarization arXiv:2608.04307v1 Announce Type: new Abstract: Text summarization is deceptively difficult. While condensing information seems straightforward, real-world enterprise summarization of support tickets, legal documents, incident reports, and more, demands strict adherence to… 23 Hugging Face Daily Papers research 24d ago TriGlue: a Biology-Inspired Generative Model for Generating Molecular Glue-Induced Ternary Complex Abstract Molecular glue degraders have emerged as a promising strategy for targeted protein degradation by inducing ternary complex formation between an E3 ubiquitin ligase and a target protein. Despite their therapeutic potential, computational design of molecular glues remains… 24 Simon Willison community 24d ago Incident Report: unsanctioned agent behaviour during cyber testing Incident Report: unsanctioned agent behaviour during cyber testing It happened again . This time it was the UK government's AI Security Institute who accidentally attacked other companies while running an evaluation with models with the safety filters turned off. From their… 37 Simon Willison community 24d ago Incident Report: unsanctioned agent behaviour during cyber testing Incident Report: unsanctioned agent behaviour during cyber testing It happened again . This time it was the UK government's AI Security Institute who accidentally attacked other companies while running an evaluation with models with the safety filters turned off. From their… 14 arXiv — Machine Learning research 25d ago Federated generative event models for tokenized electronic health records arXiv:2608.02939v1 Announce Type: new Abstract: Electronic health record foundation models are limited by institutionally siloed data and substantial performance degradation under cross-site transfer. We evaluated federated training of tokenized generative event models (GEMs)… 6 arXiv — Machine Learning research 26d ago Policy Optimality Measurement for Multi-Vehicle Decision-Making: From Extrinsic Indicators to Intrinsic Quality arXiv:2608.01133v1 Announce Type: new Abstract: Evaluating Multi-Agent Reinforcement Learning (MARL) policies in autonomous driving fundamentally relies on extrinsic statistical indicators (e.g., reward curves and success rates), which often mask intrinsic policy degradation and… 22 arXiv — Machine Learning research 27d ago PiDDM: Physics-Informed Differentiable Degradation Modeling for Lithium-Ion Battery State-of-Health Prediction arXiv:2607.29095v1 Announce Type: new Abstract: Accurate prediction of lithium-ion battery state of health (SOH) is essential for reliable energy storage operation. However, purely data-driven models may generalize poorly across cycling protocols and produce physically… 27 r/MachineLearning community 27d ago Context degradation in LLMs: what the papers actually show, and the habits I built for long analysis sessions [R]   submitted by   /u/usernamehere93 [link]   [comments] 23 TechCrunch — AI news-outlet 29d ago OpenAI reportedly finds evidence that more of its agents ran amok OpenAI has reportedly found evidence of additional agent misbehavior as it looks into the incident that occurred with Hugging Face. 24 arXiv — Machine Learning research 1mo ago Event-Structured Physics-Informed Neural Networks for Differentiable Critical Clearing Boundaries arXiv:2607.27681v1 Announce Type: new Abstract: Transient-stability assessment determines whether a power system can recover after a disturbance and is therefore essential to preventing generator trips and cascading outages. A key metric is the critical clearing time (CCT),… 7 arXiv — Machine Learning research 1mo ago VESTIGE: A Knowledge-Guided Masking Strategy for Corruption-Aware Fine-Tuning of Genomic Transformers, Validated on Ancient DNA Reconstruction arXiv:2607.27712v1 Announce Type: new Abstract: Standard masked-language-model fine-tuning applies a uniform masking probability across every token position, assuming reconstruction difficulty is position-agnostic. When the degradation process is characterised and concentrated… 4 TechCrunch — AI news-outlet 1mo ago Anthropic says its own AI models breached three companies during security tests After OpenAI's models broke into Hugging Face, Anthropic checked its own history and found three similar incidents 13 Simon Willison community 1mo ago Investigating three real-world incidents in our cybersecurity evaluations Investigating three real-world incidents in our cybersecurity evaluations It happened again! This is turning into something of a pattern. Last week OpenAI accidentally exploited Hugging Face when one of their frontier models broke out of a sandboxed container and hacked into… 10 Page 1 of 4 · 158 articles Older →