News / #outage Tag Outages 179 articles archived under #outage · RSS Sign in to follow Don't Worry About the Vase community 1h ago OpenAI and the Wiki Incident I did not expect to be back here so soon with more OpenAI agent swarm coverage. 37 TechCrunch — AI news-outlet 1d ago OpenAI confirms ‘wiki incident,’ says it’s ‘working on a framework’ for more disclosure OpenAI acknowledged its role in a recently reported incident where AI agents took over a German wiki forum. 14 The Information — AI news-outlet 1d ago OpenAI Pledges New Rules for Reporting Troubling Behavior by Its AI Agents OpenAI acknowledged that its AI agents posted messages on external wiki websites earlier this year, saying it is developing new rules for disclosing such “misalignment” incidents. The company’s post on X, published just after 12 a.m. Saturday, followed an independent report ... 17 Hacker News — AI on Front Page community 1d ago AI handles incidents, engineers lose touch with their systems Article URL: https://www.sylvainkalache.com/blog/ai-handles-incidents-engineers-lose-touch-with-their-systems Comments URL: https://news.ycombinator.com/item?id=49574167 Points: 212 # Comments: 186 27 r/LocalLLaMA community 1d ago The OpenAI Huggingface incident from an agents POV Full credits to @artificialisabel from X!   submitted by   /u/iPingWine [link]   [comments] 37 TechCrunch — AI news-outlet 1d ago OpenAI’s rogue agents keep escaping, with no formal process to investigate them OpenAI’s latest agent swarm incident adds urgency to calls for independent investigations as researchers and lawmakers question whether AI labs should control the scope of their own safety reviews. 15 arXiv — NLP / Computation & Language research 2d ago TabScope: Question-Adaptive Scope Selection for Table Question Answering arXiv:2609.03395v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong performance on table question answering, yet their accuracy often degrades as table size increases. We find that this degradation is not uniform across question types.… 7 Ars Technica — AI news-outlet 3d ago Four major AI models suffer rare overlapping downtime Service interruptions hit ChatGPT, Claude, Grok, and Gemini practically simultaneously. 28 Hacker News — AI on Front Page community 3d ago Ask HN: Why were OpenAI, Claude, and Grok simultaneously down? https://status.openai.com https://status.claude.com https://status.x.ai ChatGPT outage – Resolved - https://news.ycombinator.com/item?id=49550614 (315 comments) Claude outage – Resolved - https://news.ycombinator.com/item?id=49549676 (146 comments) Grok outage -… 13 arXiv — Machine Learning research 3d ago OutageDiT: A Generative Foundation Model for Power Outage Forecasting and Scenario Simulation arXiv:2609.01896v1 Announce Type: new Abstract: Power-outage planning requires scenarios before an event occurs. These scenarios must represent uncertainty in magnitude, timing, and duration while preserving temporal dependence. However, severe events are rare, and data from any… 14 arXiv — Machine Learning research 3d ago Multi-Agent Retrieval-Augmented Generation for Efficient Cloud Knowledge Base Search in Telecom SNOC Environment arXiv:2609.01618v1 Announce Type: cross Abstract: Telecom Service and Network Operations Centers (SNOCs) rely on large collections of cloud documents, including Standard Operating Procedures (SOPs), vendor technical manuals, incident reports, and configuration guides, to… 9 r/LocalLLaMA community 4d ago Confirmed bolting Q8 NGram into IQ4 Qwen no speed degradation This came from another thread or comment. I forgot exactly where, but the basic idea was to replace the 51B N-gram layer in Qwen 3.8 Next with a much higher precision version. Someone running a 5090 replaced the N-gram portion of their Qwen 3.8 UD Q4 model with BF16. Since I'm… 33 The Algorithmic Bridge news-outlet 4d ago The AI Industry Has a Really Dark Secret You Should Know About Review of and thoughts on the Hugging Face incident 9 arXiv — Machine Learning research 4d ago HBQ: Hierarchical Scaling Block Quantization with Hardware-Efficiency-Aware Design for Accurate LLM Inference arXiv:2609.00450v1 Announce Type: new Abstract: Block Quantization (BQ) is a promising approach for efficient deployment of large language models (LLMs), enabling low-precision computation with controlled accuracy degradation. Compared to scalar weight-only quantization (WoQ),… 8 TechCrunch — AI news-outlet 5d ago Sequoia-incubated Empirik launches with $21M to predict outages before they happen The startup wants to do for IT infrastructure what Cursor did for software engineering. 17 arXiv — Machine Learning research 5d ago HalluPrism: When Multimodal Uncertainty Should Diagnose, Not Decide arXiv:2608.29193v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) can assign similar confidence to answers that fail for different reasons. We propose HalluPrism, a behavioral diagnostic that re-runs an answer after visual degradation, blank-image… 23 MIT Technology Review — AI news-outlet 6d ago Hugging Face hack could indicate cultural issues at OpenAI This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. By now you’ve probably heard about last month’s major AI security incident, in which OpenAI agents escaped their sandbox and hacked into the… 29 Marcus on AI community 6d ago Dwarkesh Patels’s wildly popular but dangerously misleading account of the OpenAI Hugging Face incident When “plain English” isn’t a good thing. 10 arXiv — Machine Learning research 6d ago A Method for Layer Bit-Width Allocation in LLM Quantization via Performance Maximization Under a Quality-Degradation Constraint arXiv:2608.28003v1 Announce Type: new Abstract: This paper proposes a layer bit allocation method for Gemma-3-1B, formulating the problem as performance maximization (latency decrease) given a degradation budget constraint (allowable level of generation quality loss). This… 30 arXiv — NLP / Computation & Language research 6d ago DisCTI: Who Needs to Know Timely? Automated Sector-Aware Cyber Threat Intelligence Dissemination arXiv:2608.27967v1 Announce Type: cross Abstract: The timely dissemination of cyber threat intelligence (CTI) is critical for organizations to mount swift and effective incident response. When valid CTI is delivered to the right sector at the right time, identical attacks can… 5 One Useful Thing (Ethan Mollick) community 6d ago Agency and Agents From the Hugging Face Incident to Twilight Factories 15 Marcus on AI community 9d ago 5 lessons from the OpenAI / Hugging Face incident Did OpenAI really do the best they could? 8 arXiv — Machine Learning research 9d ago Beyond Capability Benchmarks: Learning Operational Fingerprints of LLM Cloud Services from Production Incident Metadata arXiv:2608.26332v1 Announce Type: new Abstract: Managed LLM services are now part of real production systems, but model selection and service planning still rely heavily on capability benchmarks that reveal little about operational behavior after deployment. We present… 26 arXiv — NLP / Computation & Language research 9d ago When Is Noise Response Universal? Tokenization as the Hidden Variable in Language Models arXiv:2608.26319v1 Announce Type: new Abstract: The performance of textual neural models often degrades when their inputs are corrupted by noise such as typos, OCR errors, or dropped words. We study the degradation rate across neural models, both sentence embeddings and… 18 TechCrunch — AI news-outlet 10d ago Here’s all the times AI has gone rogue and hacked other companies A recap of all the incidents involving LLMs made by Anthropic, Meta, and OpenAI, which went rogue and attacked real companies and individuals on the internet. 32 arXiv — Machine Learning research 10d ago FedQoS: Federated QoS-Risk Learning for Heterogeneous Indoor-Outdoor Access Selection arXiv:2608.25496v1 Announce Type: new Abstract: Reliable access selection in dynamic and heterogeneous indoor-outdoor environments is challenging because instantaneous radio measurements alone cannot capture future QoS degradation caused by mobility, blockage, traffic load, and… 33 Latent.Space news-outlet 10d ago [AINews] NVIDIA buys HuggingFace for $13B, as OpenAI publishes their HF incident retro Open Source wins! 10 Hacker News — AI on Front Page community 11d ago GitHub Outage Tracker: Is GitHub Cooked? Article URL: https://isgithubcooked.com/ Comments URL: https://news.ycombinator.com/item?id=49454728 Points: 206 # Comments: 131 9 Hacker News — AI on Front Page community 11d ago The Hugging Face incident and the road ahead Article URL: https://openai.com/index/hugging-face-incident-and-the-road-ahead/ Comments URL: https://news.ycombinator.com/item?id=49454314 Points: 214 # Comments: 258 38 TechCrunch — AI news-outlet 11d ago OpenAI releases its official report on the Hugging Face breach The report, which spans several discrete cybersecurity compromises, is the most complete accounting of the incident to date. 7 Hacker News — AI on Front Page community 11d ago Disruption with Some GitHub Services Article URL: https://www.githubstatus.com/incidents/hcbtzksccj2f Comments URL: https://news.ycombinator.com/item?id=49450722 Points: 224 # Comments: 134 36 arXiv — Machine Learning research 11d ago Data Leakage Inflates Generalizability of Power Outage Prediction Models arXiv:2608.24665v1 Announce Type: new Abstract: Power outage prediction models are increasingly used in assessments of climate-driven infrastructure risk, yet current evaluation practices obscure whether these models generalize to the novel conditions such applications require.… 6 OpenAI official-blog 11d ago The Hugging Face incident and the road ahead OpenAI shares findings from the Hugging Face security incident and the steps we’re taking to strengthen AI model security, monitoring, and alignment. 21 arXiv — NLP / Computation & Language research 12d ago Evidence-State Reliability Under Controlled Degradation: Parser-Validity Divergence in a Multi-Stage LLM Pipeline arXiv:2608.21559v1 Announce Type: new Abstract: Multi-stage LLM pipelines can remain structurally valid even when evidence available to downstream stages becomes incomplete, compressed, or conflicting. This paper introduces and operationalizes Evidence-State Reliability (ESR),… 6 arXiv — NLP / Computation & Language research 12d ago Improving Few-Step Language Flows with Untied Self-Conditioning arXiv:2608.22244v1 Announce Type: new Abstract: Flow-matching language models refine all token positions in parallel and can trade sampling steps for latency, yet generation quality still degrades sharply with few sampling steps. We trace a source of this degradation to a… 12 Hacker News — AI on Front Page community 17d ago The August 17 outage Article URL: https://github.blog/news-insights/company-news/the-august-17-outage-and-the-work-ahead/ Comments URL: https://news.ycombinator.com/item?id=49378957 Points: 499 # Comments: 568 15 r/LocalLLaMA community 17d ago Getting better at coding doesn't make a model better at everything else A majority of users in this sub use LLMs for coding/agentic tasks and I see why a lot of value is put into them but many try to say "Well coding has improved therefore it can just use tool calling and/or just look up what the user needs if there's a degradation for general… 37 r/LocalLLaMA community 17d ago Qwen3.8-23B-Mini-Me: A Depth-Pruned Qwen3.8-27B (to ~22.7BB) I've been working on a depth pruning approach and decided to try it out on the new Qwen3.8-27B model. I managed to get the model down to about 22.7B params without severe reasoning degradation. No fine-tuning was done, just strategic removal of layers. It's been working well for… 13 arXiv — NLP / Computation & Language research 18d ago AISA: AI Safety Assistant Framework for Continuous Improvement of Highway Construction arXiv:2608.17184v1 Announce Type: new Abstract: Job Safety Analysis (JSA) and pre-task planning can benefit from prior incident records, yet historical accident data is often stored as unstructured narratives that are difficult to consult at the point of planning. A novel… 15 Vercel — AI dev-tools 19d ago $1 million hacker challenge for Vercel Sandbox Agents need to run untrusted code, and the microVM has become the standard way to do it: a dedicated guest kernel per workload, isolated from the host and from every other workload on the same machine. But recent security research and real-world incidents have revealed that… 26 arXiv — Machine Learning research 19d ago Early Cycle Charge Trajectory Generative Prediction and Full Life Cycle Health Management of Iron-Chromium Flow Batteries Based on FlowBD-E1 arXiv:2608.14637v1 Announce Type: new Abstract: Long-duration stationary energy storage requires batteries whose degradation can be detected before substantial capacity loss has accumulated. Iron-chromium redox flow batteries are attractive for this role because they use… 29 arXiv — Machine Learning research 19d ago Real-Time State-of-Health Estimation and Online Degradation Prognosis from Partial Battery Discharge Using Physics-Informed Neural Networks arXiv:2608.14764v1 Announce Type: new Abstract: With the increasing integration of renewable energy sources, energy storage systems have become essential, making the accurate estimation of their State of Health (SOH) and degradation behavior critical. In this work, we propose a… 9 arXiv — NLP / Computation & Language research 19d ago Why Vision Fails as a Universal Bridge: Rectifying Modality Asynchrony in Multilingual MLLMs arXiv:2608.15085v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) exhibit substantial performance degradation in non-English visual reasoning, despite the strong multilingual competence of their text-only backbones. While mechanistic evidence from… 38 Hugging Face Daily Papers research 19d ago Ventor-QTest: Threat-Model-Driven Verification of Vendor-Hosted LLM APIs Abstract Ventor-QTest audits hosted open-weight model APIs via repeated and long-sequence black-box probes, measuring average and extreme fidelity loss to detect degradation in long-horizon agentic performance. Generated by thinkingmachines/Inkling-Small As large language models… 35 Hacker News — AI on Front Page community 20d ago Incident with Github.com Article URL: https://www.githubstatus.com/incidents/zkxwbgr0cnmx Comments URL: https://news.ycombinator.com/item?id=49330684 Points: 427 # Comments: 347 10 arXiv — Machine Learning research 24d ago GCPO: Diagnosing and Constraining Subspace Geometry in Rollout RL for LLMs arXiv:2608.11674v1 Announce Type: new Abstract: On-policy rollout methods such as GRPO are central to post-training of large language models, yet they frequently suffer from training instabilities, cross-task capability degradation, and response-length inflation. Although prior… 34 arXiv — NLP / Computation & Language research 24d ago Language-Conditional Dequantization: Recovering What Quantization Steals from Non-English Languages arXiv:2608.11786v1 Announce Type: new Abstract: Aggressive quantization disproportionately harms multilingual capability: in the sub-4B INT3 GPTQ regime, we measure 2-4x larger perplexity degradation on non-English languages than on English. We propose Language-Conditional… 28 arXiv — Machine Learning research 25d ago SQuaT: Self-Supervised Knowledge Distillation via Student-Aware Quantized Teacher Features arXiv:2608.10709v1 Announce Type: new Abstract: Quantization-Aware Training (QAT) enables the deployment of quantized models with minimal accuracy degradation. However, in practical scenarios, training labels are often unavailable due to privacy, copyright, or cost constraints.… 5 arXiv — NLP / Computation & Language research 25d ago The Multilingual Quantization Tax: Structural Collapse and Typological Fragility in Edge SLMs arXiv:2608.09941v1 Announce Type: new Abstract: While 4-bit weight quantization is critical for deploying Small Language Models (SLMs) on edge devices, evaluations of the resulting performance degradation-the quantization tax-remain overwhelmingly English-centric. We present a… 30 arXiv — Machine Learning research 27d ago MiCoPro: End-to-End Mixed Precision HW/SW Co-design with HW-aware Proxy Model arXiv:2608.06916v1 Announce Type: new Abstract: Quantized Neural Networks~(QNN) with low-bitwidth data have proven promising in efficient storage and computation on edge devices. To mitigate accuracy degradation while maximizing speedup, layer-wise mixed-precision… 31 Page 1 of 4 · 179 articles Older →