News / #hardware Tag Hardware 500 articles archived under #hardware · RSS Sign in to follow TechCrunch — AI news-outlet 10d ago Google, Nvidia and Anthropic want Emerald AI to find space on the grid for more data centers A new coalition that includes Google, Nvidia, Anthropic and Emerald AI wants to find 100 GW of grid capacity for new data centers. 22 TechCrunch — AI news-outlet 11d ago Al Gore says the real AI risk isn’t data centers — it’s what industry leaders are warning about In an interview with TechCrunch, Al Gore suggested he isn't losing sleep over AI data center emissions — he's more worried about the AI industry's own warnings about where the technology is headed. 38 NVIDIA Developer Blog official-blog 11d ago TensorRT Edge-LLM Completes the MLPerf Edge Agentic Benchmark 6.4x Faster on Jetson AGX Thor AI agents are moving from cloud data centers to vehicles, robots, and other edge devices. Unlike a chatbot that answers a single prompt, an agent works through... 14 MIT Technology Review — AI news-outlet 11d ago Building the materials foundation for AI The AI boom is becoming a materials challenge. As AI pushes computing into new territory, the materials behind that infrastructure are becoming just as crucial as the algorithms running on it. Semiconductors and data centers are approaching physical limits around performance,… 20 arXiv — Machine Learning research 12d ago Differentially Private Semantic Plans for Aggregate Insight Generation arXiv:2609.16283v1 Announce Type: new Abstract: \texttt{URANIA} provides end-to-end differential privacy (DP) for summaries of data-dependent clusters. However, its cluster--keyword release does not directly provide collection-wide aggregates for semantic concepts defined… 33 arXiv — Machine Learning research 12d ago OPEN-1B: A Fully Auditable Training Run arXiv:2609.17380v1 Announce Type: new Abstract: Open-source language models have a reproducibility problem. Despite releasing weights, training data, and recipes, none of them are provably reproducible due to the non-associativity of floating-point arithmetic. Deep learning… 19 r/LocalLLaMA community 12d ago Open Source Appreciation Post It's late at night in the lab, I've been working on a basic script for a virology project, and holy hell the safeguards have been pissing me off. Mirroring detectEVE data over rsync to my laptop by making a zip file first? No no no, great safety mogul DARIO demands there be NO… 18 TechCrunch — AI news-outlet 12d ago The AI data center boom is colliding with cities scarred by big industry National outcry against data center construction has spread to Philadelphia, where officials suggested possible construction in a neighborhood already impacted by a now-defunct oil refinery. 13 TechCrunch — AI news-outlet 12d ago US data centers could consume more natural gas than Germany and Japan combined by 2035 The AI frenzy could push U.S. data centers to become one of the largest consumers of natural gas in the world. 29 NVIDIA Developer Blog official-blog 12d ago How NVIDIA NVLink 6 Delivers Multi-Layer Resiliency for AI Factories For operators of large-scale AI factories, maximizing continuous output is essential for productivity. In massive-scale AI training, every GPU in the cluster... 37 r/MachineLearning community 12d ago Suggestions from all of you guys is needed, please respond this text [D] Hey guys, I have been posting in this subreddit for a long time. I have been asking to give me sugggesitons on my personal career, on how to move ahead in this AI replacement era. I have been touching the basics, trying to make sure I do not miss any fundamentals. I started with… 11 r/MachineLearning community 13d ago How much work in progress can a workshop submission be [R] Hi, let's suppose I am working on an algorithm that uses principles x to solve problems A and B. I already implemented a very basic algorithm that used principle "x mini" to just solve problem A, ran experiments, but have not yet implemented the full one to solve A and B. I must… 26 r/LocalLLaMA community 13d ago Harness: Am I doing something wrong? Or are my expectations unreasonable I have tried using qwen 3.6 and 3.8 q5 64k context and I'm just not getting the experience everyone here seems to get. I'm using open code and it will loop, forget what it's doing, and mess up the ui. All I am making is a simple web app that is basically a glorified text turn… 4 arXiv — Machine Learning research 13d ago SPICE: Simple Polysemantic Feature Interpretation via Clustering-based Explanation arXiv:2609.13198v1 Announce Type: new Abstract: One of the pivotal recent challenges in neural network interpretability is polysemanticity, where a single neuron is activated by multiple, often unrelated concepts, hindering clear functional understanding. Although prior work has… 29 arXiv — Machine Learning research 13d ago Specification Oracles arXiv:2609.13415v1 Announce Type: new Abstract: Specifications face a basic tradeoff: leave details out, and important questions go unanswered; record every detail separately, and the specification becomes large and prolix. We investigate whether a language model can serve as a… 6 arXiv — Machine Learning research 13d ago JumpStart Your Policy Learning with Lessons from 160,000 Training Runs arXiv:2609.13730v1 Announce Type: new Abstract: Reliable progress in offline policy learning depends on careful reporting, well-tuned baselines, and evaluation across diverse conditions. Prior work has shown that results can be sensitive to reporting choices, hyperparameter… 10 r/MachineLearning community 13d ago RSI is not happening [R] A new paper (I'm not a coauthor BTW -- I just found it interesting) argues, basically, that RSI is not on the horizon, because current (at the time the study was done) agents cannot do open-ended ML research. Specifically, they took some accepted, but unpublished papers from… 17 The Information — AI news-outlet 13d ago Refounding America: The Tax Code Needs to Change in the Age of AI Vinod Khosla is founder of Khosla Ventures and has spent the last 40 years in the business of innovation. He was the co-founder of Sun Microsystems. Khosla Ventures is an investor in OpenAI. Major technological revolutions should force a renegotiation of society’s basic economic… 8 r/MachineLearning community 13d ago [P] Wine synthesis using VAE [P] I have created a VAE model using PyTorch on White Wine dataset. Basically, the main goal is to discover a brand-new white wine recipe. It puts all the wines into a latent space, finds the best part where higher bands are located, and then it makes 100 steps with a step size of… 36 arXiv — Machine Learning research 14d ago Temporal Recurrence Favors Fewer Layers arXiv:2609.12531v1 Announce Type: new Abstract: In streaming tasks, recurrent models can carry latent computation across time, allowing each update to build on representations produced earlier. This raises a basic question: once temporal recurrence provides sequential… 18 arXiv — Machine Learning research 14d ago Clustering-Based Balanced Sampling and Allocation with Data Parallelism for High-Performance Fine-Tuning arXiv:2609.12584v1 Announce Type: new Abstract: Instruction-tuning datasets for large language models (LLMs) are often large, redundant, and imbalanced, limiting efficient adaptation. Naive large-batch fine-tuning repeatedly includes overrepresented sample groups while weakly… 30 arXiv — Machine Learning research 14d ago MCRL2: Multi-resource Cross-attention-based Representation Learning-augmented Reinforcement Learning for Cloud Microservice Scheduling arXiv:2609.13048v1 Announce Type: new Abstract: Efficient microservice scheduling is crucial for maintaining load balance across nodes in data centers and ensuring high quality of service. However, achieving this in practice remains challenging due to dynamic resource imbalance… 4 arXiv — NLP / Computation & Language research 14d ago From Bench-to-Bedside: A Review of Clinical Trials in Drug Discovery and Development arXiv:2412.09378v4 Announce Type: replace-cross Abstract: Clinical trials bridge basic research and clinical application, serving as essential steps in drug development. This review examines clinical trial phases (Phase I [safety assessment], Phase II [efficacy evaluation],… 38 r/LocalLLaMA community 14d ago Grandma harness for GPU challenged? OK I have been trying to get most of my local Qwen 3.8 27b running on my "Grandma's GPU cluster" of 2xP40. I was able to make it produce reasonably usable speeds of like up to 45 tg/s and 450 prefill with fresh context, falling to 120-ish prefill on 150K+ ctx and 12-16 tg/s.… 11 r/LocalLLaMA community 14d ago Talk me out of buying a 3rd Spark Does anyone think the gurus on the DGX Spark forum are going to figure out how to magically fit DeepSeek 4.1 Flash on a 2x cluster, or is it only possible on 3 or 4 Sparks?   submitted by   /u/Porespellar [link]   [comments] 34 r/LocalLLaMA community 14d ago Migration from Claude Code to a private local harness. Questions. I'll start by saying I'm not talking about the models themselves, I'm aware that I can't come close to something like Fable's intelligence locally. Just wanted to get that out of the way. Basically. Over the last year I've gotten quite comfortable with claude code, and it seems… 20 r/LocalLLaMA community 14d ago Qwen 3.8 27B UD-IQ4_XS even faster on 16GB CUDA This is an evolution on top of Raymond's KV cache streaming fork - all credits to what enabled this goes to him. The basic idea behind what he enabled was a pool of memory in VRAM that is used differently depending on the phase (prompt processing or decoding) and when total used… 18 The Information — AI news-outlet 14d ago Why Amazon and Microsoft Are Taking Communities’ Side Against Utilities Amazon, Microsoft, Oracle and other AI data center developers are sweetening financial offers to municipalities and regulators to gain approval for new facilities. They’re also getting smarter about standing up for consumers against utilities that have proposed to make the… 14 r/LocalLLaMA community 15d ago Benchmark your custom Pi tools A few people here mentioned interest in a way to test their custom Pi setups, so I figured I’d drop this here: RoastMyHarness The basic idea is a small engine that sets up an environment to run DeepSWE benchmark tasks using bare Pi as a control and a variant of your choice, your… 32 Hacker News — AI on Front Page community 16d ago The EPA is planning to scrap public review rules for data center pollution Article URL: https://capitalbnews.org/data-centers-permit-rules-epa/ Comments URL: https://news.ycombinator.com/item?id=49662672 Points: 235 # Comments: 159 28 arXiv — Machine Learning research 17d ago Phase-Decoupled, Model-Calibrated Power Control for Disaggregated LLM Serving arXiv:2609.11133v1 Announce Type: new Abstract: Datacenter GPU power is the binding constraint on LLM serving capacity, and production serving has shifted to prefill/decode (PD) disaggregation. Deploying NVIDIA's Max-Q inference profile on a disaggregated B200 system, we found… 18 arXiv — Machine Learning research 17d ago Hierarchical Clustering Can Jointly Satisfy Richness, Consistency, and Scale Invariance arXiv:2609.11173v1 Announce Type: new Abstract: Despite its ubiquity, clustering lacks a universally accepted definition of what is a cluster. Kleinberg's Impossibility Theorem formalizes this difficulty by showing that no flat clustering method can simultaneously satisfy three… 18 arXiv — Machine Learning research 17d ago Meta-Learning for Data-Efficient Plant Growth Estimation via Vision Transformers and Fuzzy Clustering arXiv:2609.10749v1 Announce Type: cross Abstract: Accurate plant growth estimation is essential for greenhouse monitoring, yet obtaining labeled data remains costly and time-consuming. To address this, we propose a few-shot regression framework that combines Vision Transformer… 16 arXiv — NLP / Computation & Language research 17d ago Rethinking Verbalized Confidence for LLM-as-a-Judge: A Compatibility Shift on Post-2025 Proprietary Models arXiv:2609.10996v1 Announce Type: new Abstract: Verbalized confidence, long dismissed as overconfident, coarse, and prone to round-number clustering, is now the more robust soft-scoring mechanism for LLM-as-a-Judge on top-tier proprietary models. Across SummEval, AggreFact, and… 30 The Information — AI news-outlet 17d ago Microsoft, Hurt By Server Shortage, Aims to Triple Cloud Capacity by 2032 Microsoft is planning to more than triple the size of its Azure cloud unit’s data center capacity to over 38 gigawatts by 2032, up from 12 gigawatts today, Bloomberg reported Thursday. A gigawatt of capacity would be able to power a city the size of San Francisco. The planned… 16 The Information — AI news-outlet 17d ago Oracle Reports 30% Topline Growth for August Quarter Oracle reported 30% growth in revenue of $19.3 billion for the quarter ending Aug. 31, a slightly better growth rate than the company projected in June, even as it burned $5 billion to expand the data centers that are driving its accelerating topline growth. The software and… 20 The Information — AI news-outlet 17d ago SpaceX Overhauls Data Center Build-Out, Potentially Slowing Expansion Elon Musk is famous for his “move fast” management philosophy, which he demonstrated most starkly when he built new data centers for his AI startup in record time two years ago. But a new team of rocket engineers Musk recently installed to run his data centers is taking a very… 21 The Information — AI news-outlet 17d ago Nvidia Deepens Partnership With AI Chip Startup D-Matrix Chip startup d-Matrix said Thursday it plans to use Nvidia’s networking hardware to connect its chips for running AI models to each other and to Nvidia’s Vera central processing units, which are used in data centers. The startup will use Nvidia’s NVLink Fusion products,… 18 Hugging Face Daily Papers research 18d ago AgentGrad: Intervention-guided Prompt Optimization for Multi Agent Systems Abstract AgentGrad improves multi-agent prompt optimization by identifying target agents through sequential intervention and clustering gradients semantically to avoid mixing unrelated errors. Generated by thinkingmachines/Inkling-Small Large language model (LLM)-based… 28 MIT Technology Review — AI news-outlet 18d ago Powering AI is an architecture problem On July 22, 2026, a transmission line fault in Ashburn, Virginia—the heart of the world’s largest data center cluster—knocked more than 3 gigawatts of load off the grid in seconds. And it wasn’t the first time. Two years earlier, a single failed surge arrester… 36 arXiv — Machine Learning research 18d ago Granular-Ball Quantum Clustering for Resource-Efficient and Robust Learning arXiv:2609.06016v1 Announce Type: new Abstract: Quantum clustering aims to exploit quantum feature representations to uncover complex data structures beyond conventional Euclidean geometry. Yet this sample-level kernel construction requires O(n^2) quantum circuit executions for… 20 arXiv — Machine Learning research 18d ago Sparse Incident-Cluster Learning for 12-hour Port Flood Pre-warning in Digital-Twin Analytics arXiv:2609.06109v1 Announce Type: new Abstract: Port flood digital twins require analytics that warn operators before disruption, but official warning incidents are often few and adjacent observations are temporally dependent. Row-level classification can therefore overstate… 29 arXiv — Machine Learning research 18d ago Parameterized and Streaming Algorithms for Euclidean Fair $k$-Center Clustering arXiv:2609.06384v1 Announce Type: new Abstract: Motivated by the growing importance of fairness in machine learning, fair $k$-center clustering has attracted considerable research attention as a fundamental problem. In this problem, a dataset is partitioned into $m$ disjoint… 8 arXiv — Machine Learning research 18d ago Sector-Mean: Deterministic Initialization of K-Means Centroids via Angular Sector Partitioning arXiv:2609.06468v1 Announce Type: new Abstract: K-Means is one of the most widely used clustering algorithms, but its susceptibility to initial centroid selection remains a primary bottleneck for its convergence speed and clustering accuracy. This paper proposes Sector-Mean… 35 arXiv — NLP / Computation & Language research 18d ago Can Foundation Models Moderate Online Content? Evaluating Instruction- vs. Example-Driven Policy Operationalization arXiv:2609.10410v1 Announce Type: new Abstract: The growing complexity of content moderation policies presents a critical challenge for their consistent operationalization. While foundation models possess the basic capabilities needed to confront this challenge, whether they can… 22 TechCrunch — AI news-outlet 18d ago Massachusetts hits data centers with new clean power rules Massachusetts has become the third state in as many months to slap new restrictions on data center development. 4 r/LocalLLaMA community 18d ago Adding emotion control tags to Qwen3-TTS I fine-tuned Qwen3-TTS with inline transcript control tags rather in lieu of a separate instruction parameter and thought the community here might be interested in the result. https://huggingface.co/SpragAI/qwen3-tts-emotion-tags The basic premise was to use Qwen's CustomVoice… 14 The Information — AI news-outlet 19d ago OpenAI Works With Samsung on Next-Generation Chips OpenAI said it is working with Samsung Electronics to develop next-generation chips, deepening a collaboration with the South Korean technology giant that already spans data centers and enterprise sales, Reuters reported. “One of the areas where we have made the most progress… 6 The Information — AI news-outlet 19d ago OpenAI Works With Samsung on Next-Generation Chips OpenAI said it is working with Samsung Electronics to develop next-generation chips, deepening a collaboration with the South Korean technology giant that already spans data centers and enterprise sales, Reuters reported. “One of the areas where we have made the most progress… 28 r/MachineLearning community 19d ago Is anyone working on wave-superposition-based pattern recognition instead of neural-network weights? [R] I’ve been thinking about an alternative way of doing low-level AI perception, and I’m curious whether anyone here is already working on something similar. The basic idea is to use waves and physical superposition/interference as the computational substrate , instead of doing… 5 Page 2 of 10 · 500 articles ← Newer Older →