News / #outage Tag Outages 212 articles archived under #outage · RSS Sign in to follow arXiv — NLP / Computation & Language research 7h ago Estimating and Orthogonalizing Unknown Pre-training Gradients for Continual Fine-tuning of Large Language Models arXiv:2609.30935v1 Announce Type: new Abstract: Continual fine-tuning is essential for large language models (LLMs) to dynamically adapt to real-world environments, yet it inevitably suffers from catastrophic forgetting, particularly the performance degradation of previous tasks… 19 Marcus on AI community 1d ago BREAKING: AI agent incident toll has risen to tens of thousands Meanwhile, the US government appears to be paralyzed 37 r/LocalLLaMA community 2d ago Qwen3.8-27B: Using KV Cache Transplants to Boost Output Quality Since my last post , I've been thinking about different options for dynamic performance degradation, trying to squeeze as much high-quality inference out of my GPU as I can. Over the weekend I read this really interesting paper: Cache-to-Cache: Direct Semantic Communication… 18 TechCrunch — AI news-outlet 3d ago Australia to investigate if OpenAI hack of government health website broke the law The incident is the first known breach to affect a government agency, and Australia's prime minister has vowed to hold OpenAI accountable. 9 MIT Technology Review — AI news-outlet 6d ago Don’t be fooled by this summer of AI hype It’s been a busy few months for AI hype. At the end of April, Anthropic claimed that its model Claude Mythos is better at finding software vulnerabilities than most security experts. Then we had the OpenAI–Hugging Face hacking incident, after which Anthropic (proudly) and Meta… 19 arXiv — NLP / Computation & Language research 6d ago Do Student LLMs Inherit OOD Robustness? Invariance-Weighted Distillation for Reliable Knowledge Transfer arXiv:2609.22566v1 Announce Type: new Abstract: Knowledge distillation (KD) aims to compress high-performance teacher LLMs into lightweight students. However, distilled students often exhibit substantial performance degradation in out-of-distribution (OOD) settings, a critical… 27 The Information — AI news-outlet 6d ago OpenAI and Anthropic Neared Deal to Stress-Test Each Other’s AI OpenAI is rethinking a range of safety strategies as it responds to fears from employees and others about the dangers its AI poses. One solution could lie in the recent past. Even before the spate of cybersecurity incidents involving OpenAI’s technology and the dire warnings… 24 arXiv — Machine Learning research 7d ago Joint Remaining Useful Life Prediction and Capacity Estimation of Lithium-Ion Batteries Using Partial-Charging Data arXiv:2609.21932v1 Announce Type: new Abstract: Joint remaining useful life (RUL) prediction and capacity estimation require representations of both gradual degradation and recent battery behavior. This paper presents a cross-expert framework using partial-charging measurements… 28 arXiv — NLP / Computation & Language research 7d ago Accelerating Dense LLMs via L0-regularized Mixture-of-Experts arXiv:2609.21672v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance but suffer from slow and costly inference. Existing acceleration methods often lead to noticeable performance degradation, while Mixture-of-Experts (MoE) models require… 31 Don't Worry About the Vase community 8d ago Anthropic Looks At Some Of Its Alignment Problems Anthropic has given us its assessment of four ‘recent cybersecurity incidents’ involving Claude that happened during cybersecurity evaluations, three of which were previously known. 32 Ars Technica — AI news-outlet 10d ago Covert uploads and megalomania: OpenAI details new "misaligned" agent incidents Model maker commits to new framework for reporting misaligned models. 26 The Information — AI news-outlet 11d ago OpenAI Discloses More Safety Incidents and Adopts New Reporting Framework OpenAI released a new framework on Wednesday for how it aims to report unsafe or concerning behavior in its models, and disclosed six incidents of such behavior it had observed in the past six months. The framework comes after multiple employee warnings and hacking incidents… 6 Hacker News — AI on Front Page community 12d ago Salesforce Global Outage Article URL: https://status.salesforce.com/products/all Comments URL: https://news.ycombinator.com/item?id=49724488 Points: 221 # Comments: 126 18 arXiv — NLP / Computation & Language research 12d ago Few-Shot Degradation Is Not What It Seems: Behavioral Evidence, Representation Analysis, and a Random-Text Control Across 12 Models, 2 Tasks, and 2 Architectures arXiv:2609.15990v1 Announce Type: new Abstract: Few-shot prompting sometimes degrades language models instead of helping them, but why this happens is unknown. We evaluate 12 open-weight models on two Ukrainian tasks news classification and legal case outcome prediction and find… 32 arXiv — NLP / Computation & Language research 12d ago What Breaks Under Pruning in Smart Homes, and When? Evaluating LLM Degradation Across Architectures and Task Complexity arXiv:2609.17515v1 Announce Type: new Abstract: Pruning can reduce the deployment cost of large language models (LLMs), but its impact on context-grounded tool calling remains poorly understood. We systematically study pruning-induced degradation in smart-home tool calling… 14 arXiv — Machine Learning research 13d ago A Machine Learning Framework for Fault Detection, Isolation, and Severity Prediction of Autonomous VTOL Aircraft arXiv:2609.14180v1 Announce Type: new Abstract: Fault detection in autonomous VTOL aircraft is critical because even minor component degradations can rapidly destabilize multirotor vehicles operating in complex, safety-critical environments, motivating robust fault detection and… 28 arXiv — NLP / Computation & Language research 13d ago Forty Shades of Blue: Quality-Diversity Alignment via Mode-Conditioned Reinforcement Learning arXiv:2609.14896v1 Announce Type: new Abstract: A notable byproduct of LLM alignment training is mode collapse: the progressive loss of output diversity that narrows a model's expressivity at inference time. This degradation is especially limiting for applications requiring… 27 The Information — AI news-outlet 16d ago OpenAI AI Swarm Hacked Software Service Months Before Hugging Face Incident A swarm of OpenAI agents conducted a cyberattack on software service RubyGems in May, months before the company’s agents hacked model platform Hugging Face, researchers at AI safety organizations Nightingale Collective and AI Futures Project found. RubyGems allows software… 36 arXiv — Machine Learning research 18d ago Sparse Incident-Cluster Learning for 12-hour Port Flood Pre-warning in Digital-Twin Analytics arXiv:2609.06109v1 Announce Type: new Abstract: Port flood digital twins require analytics that warn operators before disruption, but official warning incidents are often few and adjacent observations are temporally dependent. Row-level classification can therefore overstate… 29 arXiv — Machine Learning research 18d ago FMMO: Detecting the Divergence Between Local Attribution and Global Drift arXiv:2609.06173v1 Announce Type: new Abstract: Post-deployment drift poses a critical risk to algorithmic accountability, particularly when ground truth labels are delayed and performance degradation becomes a "silent failure". While Explainable AI (XAI) is often relied upon to… 27 arXiv — Machine Learning research 18d ago On BatchNorm Forward Modes in Value-Based Reinforcement Learning arXiv:2609.06421v1 Announce Type: new Abstract: Batch normalization (BN) substantially improves sample efficiency in continuous-control actor-critic methods such as CrossQ, yet recent studies report performance degradation in discrete-action value learning on Atari. These… 27 arXiv — Machine Learning research 18d ago How Does Parameter Pruning Reshape DNN Representations? An Interaction-Driven Exploration arXiv:2609.06483v1 Announce Type: new Abstract: This study focuses on the scientific problem of understanding internal factors that govern the diverse performance degradation of deep neural networks (DNNs) when different parameters are pruned. In order to explain why pruning… 19 arXiv — NLP / Computation & Language research 18d ago When Auditors Fabricate: Batch-Size Degradation and Confident Hallucination in LLM Detection of Planted Document Contamination arXiv:2609.09696v1 Announce Type: new Abstract: Large language models are increasingly proposed as automated auditors of document quality, yet their reliability as detectors of planted errors is poorly characterised. We construct a contaminated corpus of 150 academic papers… 31 The Information — AI news-outlet 18d ago Anthropic Discloses Fourth Cybersecurity Incident Anthropic said Wednesday it had found a fourth cybersecurity incident involving its Claude models. The company disclosed the new incident, which occurred in January and involved an early version of Claude Opus 4.6, and additional information on three previously known incidents… 34 The Information — AI news-outlet 18d ago Anthropic Discloses Fourth Cybersecurity Incident Anthropic said Wednesday it had found a fourth cybersecurity incident involving its Claude models. The company disclosed the new incident, which occurred in January and involved an early version of Claude Opus 4.6, and additional information on three previously known incidents… 24 TechCrunch — AI news-outlet 18d ago Superintelligence is coming. Should we let it? AI companies have been talking about superintelligent AI like it’s inevitable, but recent safety incidents like OpenAI’s Hugging Face breach are demonstrating the potential dangers of deploying AI systems that are more… 7 TechCrunch — AI news-outlet 18d ago ControlAI’s Connor Leahy on why superintelligence is ‘not a weapon, it’s an adversary’ AI companies have been talking about superintelligent AI like it’s inevitable, but recent safety incidents like OpenAI’s Hugging Face breach are demonstrating the potential dangers of deploying AI systems that are more… 35 Hugging Face Daily Papers research 18d ago Counter-Swarm Doctrine: Containing Coordinated Agent Intrusions Abstract Agents can turn shared infrastructure into a channel for coordinated intrusion. The Hugging Face incident and a separate public-wiki investigation show why a security assessment may need evidence from several executions and the artifacts they leave behind. We argue that… 35 Smol AI News news-outlet 19d ago not much happened today **Anthropic** disclosed four cyber incidents involving **Claude** during third-party security tests, revealing failures in situational awareness and monitorability, with an independent investigation by **METR** underway. The governance debate intensified following **Jacob… 28 r/MachineLearning community 20d ago My lab found a way to migrate between embedding models with zero downtime. [R] So I've been messinga round with embedding models for a bit, and I think they are interesting enough to experiment with. They are useful for rag, especially in a localllm sense because you can ground your answers in truth. But what happens if you have a billion documents, and… 29 arXiv — NLP / Computation & Language research 21d ago Evaluating Large Language Models for Forced Outage Risk Prediction: Benefits and Comparison to Machine Learning arXiv:2609.04272v1 Announce Type: cross Abstract: This study examines the ability of large language models (LLMs) to predict the risk of weather-related forced outages in the distribution grid in a zero-shot framework, without labeled training data. The problem is formulated as… 29 Don't Worry About the Vase community 21d ago OpenAI and the Wiki Incident I did not expect to be back here so soon with more OpenAI agent swarm coverage. 37 TechCrunch — AI news-outlet 22d ago OpenAI confirms ‘wiki incident,’ says it’s ‘working on a framework’ for more disclosure OpenAI acknowledged its role in a recently reported incident where AI agents took over a German wiki forum. 14 The Information — AI news-outlet 22d ago OpenAI Pledges New Rules for Reporting Troubling Behavior by Its AI Agents OpenAI acknowledged that its AI agents posted messages on external wiki websites earlier this year, saying it is developing new rules for disclosing such “misalignment” incidents. The company’s post on X, published just after 12 a.m. Saturday, followed an independent report ... 17 Hacker News — AI on Front Page community 23d ago AI handles incidents, engineers lose touch with their systems Article URL: https://www.sylvainkalache.com/blog/ai-handles-incidents-engineers-lose-touch-with-their-systems Comments URL: https://news.ycombinator.com/item?id=49574167 Points: 212 # Comments: 186 27 r/LocalLLaMA community 23d ago The OpenAI Huggingface incident from an agents POV Full credits to @artificialisabel from X!   submitted by   /u/iPingWine [link]   [comments] 37 Latent.Space news-outlet 23d ago [AINews] Collusion.wiki: A second undisclosed OpenAI agent swarm incident... AI News for 9/2/2026-9/3/2026. 33 TechCrunch — AI news-outlet 23d ago OpenAI’s rogue agents keep escaping, with no formal process to investigate them OpenAI’s latest agent swarm incident adds urgency to calls for independent investigations as researchers and lawmakers question whether AI labs should control the scope of their own safety reviews. 15 Smol AI News news-outlet 24d ago collusion.wiki **OpenAI** agents were found colluding via a German-language wiki/forum, exchanging **~18,000 messages** and bypassing restrictions by exploiting writable web surfaces like public wikis and CGI endpoints. The incident raised concerns about **OpenAI's** transparency and… 26 arXiv — NLP / Computation & Language research 24d ago TabScope: Question-Adaptive Scope Selection for Table Question Answering arXiv:2609.03395v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong performance on table question answering, yet their accuracy often degrades as table size increases. We find that this degradation is not uniform across question types.… 7 Ars Technica — AI news-outlet 24d ago Four major AI models suffer rare overlapping downtime Service interruptions hit ChatGPT, Claude, Grok, and Gemini practically simultaneously. 28 Hacker News — AI on Front Page community 24d ago Ask HN: Why were OpenAI, Claude, and Grok simultaneously down? https://status.openai.com https://status.claude.com https://status.x.ai ChatGPT outage – Resolved - https://news.ycombinator.com/item?id=49550614 (315 comments) Claude outage – Resolved - https://news.ycombinator.com/item?id=49549676 (146 comments) Grok outage -… 13 arXiv — Machine Learning research 25d ago OutageDiT: A Generative Foundation Model for Power Outage Forecasting and Scenario Simulation arXiv:2609.01896v1 Announce Type: new Abstract: Power-outage planning requires scenarios before an event occurs. These scenarios must represent uncertainty in magnitude, timing, and duration while preserving temporal dependence. However, severe events are rare, and data from any… 14 arXiv — Machine Learning research 25d ago Multi-Agent Retrieval-Augmented Generation for Efficient Cloud Knowledge Base Search in Telecom SNOC Environment arXiv:2609.01618v1 Announce Type: cross Abstract: Telecom Service and Network Operations Centers (SNOCs) rely on large collections of cloud documents, including Standard Operating Procedures (SOPs), vendor technical manuals, incident reports, and configuration guides, to… 9 r/LocalLLaMA community 25d ago Confirmed bolting Q8 NGram into IQ4 Qwen no speed degradation This came from another thread or comment. I forgot exactly where, but the basic idea was to replace the 51B N-gram layer in Qwen 3.8 Next with a much higher precision version. Someone running a 5090 replaced the N-gram portion of their Qwen 3.8 UD Q4 model with BF16. Since I'm… 33 The Algorithmic Bridge news-outlet 25d ago The AI Industry Has a Really Dark Secret You Should Know About Review of and thoughts on the Hugging Face incident 9 arXiv — Machine Learning research 26d ago HBQ: Hierarchical Scaling Block Quantization with Hardware-Efficiency-Aware Design for Accurate LLM Inference arXiv:2609.00450v1 Announce Type: new Abstract: Block Quantization (BQ) is a promising approach for efficient deployment of large language models (LLMs), enabling low-precision computation with controlled accuracy degradation. Compared to scalar weight-only quantization (WoQ),… 8 TechCrunch — AI news-outlet 26d ago Sequoia-incubated Empirik launches with $21M to predict outages before they happen The startup wants to do for IT infrastructure what Cursor did for software engineering. 17 arXiv — Machine Learning research 27d ago HalluPrism: When Multimodal Uncertainty Should Diagnose, Not Decide arXiv:2608.29193v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) can assign similar confidence to answers that fail for different reasons. We propose HalluPrism, a behavioral diagnostic that re-runs an answer after visual degradation, blank-image… 23 MIT Technology Review — AI news-outlet 27d ago Hugging Face hack could indicate cultural issues at OpenAI This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. By now you’ve probably heard about last month’s major AI security incident, in which OpenAI agents escaped their sandbox and hacked into the… 29 Page 1 of 5 · 212 articles Older →