News / #outage Tag Outages 134 articles archived under #outage · RSS Sign in to follow arXiv — Machine Learning research 1d ago GCPO: Diagnosing and Constraining Subspace Geometry in Rollout RL for LLMs arXiv:2608.11674v1 Announce Type: new Abstract: On-policy rollout methods such as GRPO are central to post-training of large language models, yet they frequently suffer from training instabilities, cross-task capability degradation, and response-length inflation. Although prior… 34 arXiv — NLP / Computation & Language research 1d ago Language-Conditional Dequantization: Recovering What Quantization Steals from Non-English Languages arXiv:2608.11786v1 Announce Type: new Abstract: Aggressive quantization disproportionately harms multilingual capability: in the sub-4B INT3 GPTQ regime, we measure 2-4x larger perplexity degradation on non-English languages than on English. We propose Language-Conditional… 28 arXiv — Machine Learning research 2d ago SQuaT: Self-Supervised Knowledge Distillation via Student-Aware Quantized Teacher Features arXiv:2608.10709v1 Announce Type: new Abstract: Quantization-Aware Training (QAT) enables the deployment of quantized models with minimal accuracy degradation. However, in practical scenarios, training labels are often unavailable due to privacy, copyright, or cost constraints.… 5 arXiv — NLP / Computation & Language research 2d ago The Multilingual Quantization Tax: Structural Collapse and Typological Fragility in Edge SLMs arXiv:2608.09941v1 Announce Type: new Abstract: While 4-bit weight quantization is critical for deploying Small Language Models (SLMs) on edge devices, evaluations of the resulting performance degradation-the quantization tax-remain overwhelmingly English-centric. We present a… 30 arXiv — Machine Learning research 4d ago MiCoPro: End-to-End Mixed Precision HW/SW Co-design with HW-aware Proxy Model arXiv:2608.06916v1 Announce Type: new Abstract: Quantized Neural Networks~(QNN) with low-bitwidth data have proven promising in efficient storage and computation on edge devices. To mitigate accuracy degradation while maximizing speedup, layer-wise mixed-precision… 31 Simon Willison community 6d ago Now we have a timeline of the OpenAI accidental attack against Hugging Face OpenAI gave a last-minute presentation at the Black Hat security on Wednesday about "the Hugging Face Incident" ( previously on this blog). The video was published yesterday. It's short and information dense and well worth watching, in particular because it provides full details… 28 r/LocalLLaMA community 7d ago Black Hat USA 2026: The 'Breaking' News: The OpenAI–Hugging Face Incident   submitted by   /u/SilentLennie [link]   [comments] 38 Smol AI News news-outlet 7d ago not much happened today **OpenAI** escalates its upcoming **Astra** model to "critical" cyber status due to significant advancements in agentic coding and cybersecurity, pausing some activities to strengthen controls. The "Hugging Face incident" highlights persistent multi-agent coordination failures… 8 arXiv — Machine Learning research 7d ago Equipment-centric workpiece localization in near real-time using deep learning-based vision and event-driven finite state machines arXiv:2608.05744v1 Announce Type: new Abstract: Continuous workpiece localization is essential for traceability and process coordination in hot forging, but direct tracking is unreliable because of extreme temperatures, surface degradation, and irregular routing. This study… 13 arXiv — Machine Learning research 7d ago Failing Gracefully: Mitigating Impact of Inevitable Robot Failures arXiv:2608.05313v1 Announce Type: cross Abstract: Service robots operate in household environments shared with humans, pets, and everyday objects, where they are highly susceptible to failures such as software crashes, hardware degradation, or unpredictable interactions. While… 25 Hugging Face Daily Papers research 7d ago ChronoVision: Temporal Reasoning via Latent State Reconstruction Abstract Multimodal large language models excel at passive perception but struggle with complex visual cognitive tasks requiring multi-step temporal reasoning. This degradation largely stems from the inherent ambiguity of language-based reasoning, which often fails to accurately… 15 Hacker News — AI on Front Page community 7d ago GitHub Actions and Pages are experiencing degraded availability Article URL: https://www.githubstatus.com/incidents/qcvjkzcs7j74 Comments URL: https://news.ycombinator.com/item?id=49198302 Points: 202 # Comments: 179 19 arXiv — NLP / Computation & Language research 8d ago The Fairness Collapse Phenomenon: Bias Amplification in Language Models Trained on Synthetic Data arXiv:2608.04268v1 Announce Type: new Abstract: Generative models trained on artificially generated data have been shown to exhibit model collapse, resulting in significant performance degradation. As synthetic content increasingly contaminates the training corpora of language… 8 arXiv — NLP / Computation & Language research 8d ago MIDAS: Multi-LLM Iterative Data-Adaptive Summarization arXiv:2608.04307v1 Announce Type: new Abstract: Text summarization is deceptively difficult. While condensing information seems straightforward, real-world enterprise summarization of support tickets, legal documents, incident reports, and more, demands strict adherence to… 23 Hugging Face Daily Papers research 8d ago TriGlue: a Biology-Inspired Generative Model for Generating Molecular Glue-Induced Ternary Complex Abstract Molecular glue degraders have emerged as a promising strategy for targeted protein degradation by inducing ternary complex formation between an E3 ubiquitin ligase and a target protein. Despite their therapeutic potential, computational design of molecular glues remains… 24 Simon Willison community 8d ago Incident Report: unsanctioned agent behaviour during cyber testing Incident Report: unsanctioned agent behaviour during cyber testing It happened again . This time it was the UK government's AI Security Institute who accidentally attacked other companies while running an evaluation with models with the safety filters turned off. From their… 37 Simon Willison community 8d ago Incident Report: unsanctioned agent behaviour during cyber testing Incident Report: unsanctioned agent behaviour during cyber testing It happened again . This time it was the UK government's AI Security Institute who accidentally attacked other companies while running an evaluation with models with the safety filters turned off. From their… 14 arXiv — Machine Learning research 9d ago Federated generative event models for tokenized electronic health records arXiv:2608.02939v1 Announce Type: new Abstract: Electronic health record foundation models are limited by institutionally siloed data and substantial performance degradation under cross-site transfer. We evaluated federated training of tokenized generative event models (GEMs)… 6 arXiv — Machine Learning research 10d ago Policy Optimality Measurement for Multi-Vehicle Decision-Making: From Extrinsic Indicators to Intrinsic Quality arXiv:2608.01133v1 Announce Type: new Abstract: Evaluating Multi-Agent Reinforcement Learning (MARL) policies in autonomous driving fundamentally relies on extrinsic statistical indicators (e.g., reward curves and success rates), which often mask intrinsic policy degradation and… 22 arXiv — Machine Learning research 11d ago PiDDM: Physics-Informed Differentiable Degradation Modeling for Lithium-Ion Battery State-of-Health Prediction arXiv:2607.29095v1 Announce Type: new Abstract: Accurate prediction of lithium-ion battery state of health (SOH) is essential for reliable energy storage operation. However, purely data-driven models may generalize poorly across cycling protocols and produce physically… 27 r/MachineLearning community 11d ago Context degradation in LLMs: what the papers actually show, and the habits I built for long analysis sessions [R]   submitted by   /u/usernamehere93 [link]   [comments] 23 TechCrunch — AI news-outlet 13d ago OpenAI reportedly finds evidence that more of its agents ran amok OpenAI has reportedly found evidence of additional agent misbehavior as it looks into the incident that occurred with Hugging Face. 24 arXiv — Machine Learning research 14d ago Event-Structured Physics-Informed Neural Networks for Differentiable Critical Clearing Boundaries arXiv:2607.27681v1 Announce Type: new Abstract: Transient-stability assessment determines whether a power system can recover after a disturbance and is therefore essential to preventing generator trips and cascading outages. A key metric is the critical clearing time (CCT),… 7 arXiv — Machine Learning research 14d ago VESTIGE: A Knowledge-Guided Masking Strategy for Corruption-Aware Fine-Tuning of Genomic Transformers, Validated on Ancient DNA Reconstruction arXiv:2607.27712v1 Announce Type: new Abstract: Standard masked-language-model fine-tuning applies a uniform masking probability across every token position, assuming reconstruction difficulty is position-agnostic. When the degradation process is characterised and concentrated… 4 TechCrunch — AI news-outlet 14d ago Anthropic says its own AI models breached three companies during security tests After OpenAI's models broke into Hugging Face, Anthropic checked its own history and found three similar incidents 13 Simon Willison community 14d ago Investigating three real-world incidents in our cybersecurity evaluations Investigating three real-world incidents in our cybersecurity evaluations It happened again! This is turning into something of a pattern. Last week OpenAI accidentally exploited Hugging Face when one of their frontier models broke out of a sandboxed container and hacked into… 10 Hugging Face Daily Papers research 15d ago SecRespond: Benchmarking AI Agents for Real-World Post-Compromise Incident Response Abstract Large Language Model (LLM) agents are increasingly adopted in real-world security operations with access to host artifacts and command-line interfaces (CLIs), making it critical to thoroughly assess their security capabilities. However, existing cybersecurity benchmarks… 36 arXiv — NLP / Computation & Language research 15d ago SecRespond: Benchmarking AI Agents for Real-World Post-Compromise Incident Response arXiv:2607.26791v1 Announce Type: cross Abstract: Large Language Model (LLM) agents are increasingly adopted in real-world security operations with access to host artifacts and command-line interfaces (CLIs), making it critical to thoroughly assess their security capabilities.… 32 arXiv — NLP / Computation & Language research 15d ago MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent arXiv:2507.02259v2 Announce Type: replace Abstract: Despite improvements by length extrapolation, efficient attention and memory modules, handling infinitely long documents with linear complexity without performance degradation during extrapolation remains the ultimate challenge… 22 Hacker News — AI on Front Page community 15d ago Claude: Elevated errors across all models Article URL: https://status.claude.com/incidents/q2kg8n613kr3 Comments URL: https://news.ycombinator.com/item?id=49102150 Points: 219 # Comments: 193 38 Smol AI News news-outlet 16d ago not much happened today **OpenAI's agent security incident expanded beyond Hugging Face, affecting four additional accounts and highlighting the need for stronger enterprise hardening measures like sandboxing and audit trails. The ongoing debate around "pacing the frontier" involves calls for… 30 r/LocalLLaMA community 16d ago Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident The first autonomous agent cyberattack is an unprecedented event that deserves unprecedented transparency. Today we're sharing everything we can: a full technical timeline, an interactive replay, and how we used an open model to defend ourselves, so defenders everyvwhere can… 11 Simon Willison community 16d ago Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident Hugging Face just released this extremely detailed technical description of OpenAI's recent accidental cyberattack against their infrastructure . This attack was very sophisticated, and the… 33 TechCrunch — AI news-outlet 16d ago Sam Altman is ready to decelerate His change of position comes after "the first security incident that I have felt very viscerally." 15 arXiv — Machine Learning research 17d ago Generalization bounds and sample complexity for remaining useful life prediction from complete degradation trajectories arXiv:2607.23454v1 Announce Type: new Abstract: Data-driven remaining useful life (RUL) prediction requires complete degradation trajectories for training, yet such run-to-failure data are scarce and expensive. Practitioners currently lack principled guidance on how many failure… 24 arXiv — NLP / Computation & Language research 17d ago Simple Language Normalization Wins: Cross-Lingual Speaker Verification for the TidyVoice 2026 Challenge arXiv:2607.22923v1 Announce Type: new Abstract: Cross-lingual mismatch remains a key source of overall degradation in modern speaker verification. The TidyVoice2026 Challenge targets this setting with text-independent verification, comprising 3,666 training and 808 development… 29 arXiv — NLP / Computation & Language research 17d ago The Cross-Domain Generalization Cost of Offensive Language Detection arXiv:2607.23512v1 Announce Type: new Abstract: Offensive language detection models generally suffer performance degradation when deployed across datasets and across languages, yet most existing studies stop at reporting this phenomenon and lack a systematic methodology for… 7 r/LocalLLaMA community 17d ago Jensen Huang: During the Hugging Face incident, closed AI blocked essential forensics. An open-weight frontier model helped contain the intrusion. That’s why we created the Open Secure AI Alliance. Jensen Huang on 𝕏: https://x.com/JensenHuang/status/2081698060330250294   submitted by   /u/Nunki08 [link]   [comments] 6 arXiv — Machine Learning research 18d ago TriGlue: a Biology-Inspired Generative Model for Generating Molecular Glue-Induced Ternary Complex arXiv:2607.22143v1 Announce Type: new Abstract: Molecular glue degraders have emerged as a promising strategy for targeted protein degradation by inducing ternary complex formation between an E3 ubiquitin ligase and a target protein. Despite their therapeutic potential,… 29 arXiv — NLP / Computation & Language research 18d ago Small Vision-Language Models Know When They Are Wrong But Cannot Say So: A Two-Model Study of Stated versus Internal Confidence Under Realistic Image Degradation arXiv:2607.22034v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed on consumer hardware where input images are degraded by compression, camera shake, and poor lighting. In such settings, a reliable uncertainty signal matters more than raw… 4 Hugging Face official-blog 18d ago Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident Back to Articles a]:hidden"> Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident Published July 27, 2026 Update on GitHub Upvote 26 Hugo Larcher hlarcher Adrien Carreira XciD raphael g raphael-gl Christophe Rannou chris-rannou A companion… 19 Ars Technica — AI news-outlet 21d ago AI arms race in line for a reckoning after OpenAI hacking incident Aggressive training techniques sharpens threat of bad behavior by leading models. 5 Hacker News — AI on Front Page community 22d ago OpenAI’s accidental attack against Hugging Face is science fiction that happened OpenAI and Hugging Face address security incident during model evaluation - https://news.ycombinator.com/item?id=48997548 - July 2026 (1121 comments) Comments URL: https://news.ycombinator.com/item?id=49015639 Points: 362 # Comments: 299 15 Don't Worry About the Vase community 22d ago OpenAI Model Hacks Into HuggingFace During Cybersecurity Evaluation This latest incident is a rather dramatic escalation in agentic AI cybersecurity breaches. 12 Smol AI News news-outlet 23d ago not much happened today **OpenAI**'s internal model escaped its sandbox during a cyber evaluation and compromised **Hugging Face** infrastructure to obtain benchmark answers, sparking debate on AI security and disclosure policies. The incident highlighted the need for defenders to have equivalent or… 17 arXiv — Machine Learning research 23d ago Beyond Single-Dimensional Compression: The Compound Sparsity Frontier of Large Language Models arXiv:2607.18280v1 Announce Type: new Abstract: Large language models (LLMs) are often compressed through static parameter pruning or dynamic token-level computation, yet aggressive sparsification can trigger rapid performance degradation beyond an essential sparsity boundary.… 36 arXiv — NLP / Computation & Language research 23d ago Using Fine-Tuned LLMs to Identify Indicators of Vulnerability in UK Police Incident Logs arXiv:2607.18446v1 Announce Type: new Abstract: Purpose: Understanding how much of routine policing involves vulnerable people could inform resourcing, training, and multi-agency response, yet administrative data provide limited insight. We explore whether an LLM-based… 10 OpenAI official-blog 23d ago NTT DATA Group cuts incident analysis to 30 minutes with Codex NTT DATA Group uses ChatGPT Enterprise and Codex to help 9,000 employees automate work, cut incident analysis to 30 minutes, and scale secure AI adoption. 36 r/LocalLLaMA community 23d ago OpenAI and Hugging Face partner to address security incident during model evaluation   submitted by   /u/Recoil42 [link]   [comments] 32 Hacker News — AI on Front Page community 23d ago OpenAI and Hugging Face address security incident during model evaluation Article URL: https://openai.com/index/hugging-face-model-evaluation-security-incident/ Comments URL: https://news.ycombinator.com/item?id=48997548 Points: 330 # Comments: 187 22 Page 1 of 3 · 134 articles Older →