News / #developer-tool Tag Developer Tool 500 articles archived under #developer-tool · RSS Sign in to follow r/MachineLearning community 14d ago Got scipy's KD-tree to handle inserts and deletes without rebuilding. Three things I learned [P] I built a small library called whitetree for exact Mahalanobis nearest-neighbour search on low-dimensional sensor data that keeps arriving. The idea is old. Whiten with the Cholesky factor of the covariance so Mahalanobis becomes Euclidean, then keep several scipy cKDTrees… 34 llama.cpp releases dev-tools 14d ago b10944 sycl : Fix get mem error ( #28227 ) fix for unsupport zes API optimize the code adjust the log level rm unused head files Update docs/backend/SYCL.md Co-authored-by: Titaniumtown [email protected] fix the error to detect level zero SDK/dev package, stop build after detect… 29 Hacker News — AI on Front Page community 15d ago Linux Zoom client proactively reading everything written to X11 clipboard Article URL: https://hachyderm.io/@simontatham/117201594980991062 Comments URL: https://news.ycombinator.com/item?id=49675902 Points: 208 # Comments: 63 12 TechCrunch — AI news-outlet 16d ago Kimi-maker Moonshot AI targets $2 billion in annual revenue While K3's usage figures have declined slightly in recent months, OpenRouter data currently shows as many as 300 billion tokens being generated each day by K3 models on the system. 23 arXiv — Machine Learning research 17d ago DR-LabStack: Design and Implementation of a Clinician-Facing Web System for Diabetic Retinopathy Prediction arXiv:2609.10796v1 Announce Type: new Abstract: Pretrained diabetic retinopathy (DR) prediction models differ in their input fields, serialization formats, preprocessing requirements, and output semantics. Making these models accessible through a common clinical interface… 11 arXiv — Machine Learning research 17d ago Musec: MomentUm SpEctral Clipping for Stable Muon-type Training arXiv:2609.11655v1 Announce Type: new Abstract: Muon has emerged as a highly effective optimizer for large language model training, often achieving superior convergence and performance compared with the widely adopted Adam and AdamW optimizers. Nevertheless, Muon is prone to… 8 arXiv — NLP / Computation & Language research 17d ago Cross-Lingual Clinical Annotation Projection as Constrained Text Generation: A Six-Language Study arXiv:2609.11450v1 Announce Type: new Abstract: Background: To determine whether cross-lingual clinical annotation projection can be formulated as a text-preserving, document-level generative task that produces verifiable character-level annotations for multilingual clinical… 8 arXiv — NLP / Computation & Language research 17d ago Component-Aware Differential Privacy for Federated Multilingual Speech-LLMs arXiv:2609.11762v1 Announce Type: new Abstract: Per-layer differential privacy (DP) clipping improves gradient fidelity in federated learning by allocating per-matrix clipping budgets proportional to parameter count. We show that this recipe breaks for speech large language… 13 arXiv — NLP / Computation & Language research 17d ago The widening evaluation gap in medical large language model research 2023 to 2026 arXiv:2609.11770v1 Announce Type: new Abstract: Large language models are superseded every few quarters; clinical evidence takes years. We asked whether medical research is keeping pace with the systems it evaluates. PubMed returned 11,628 records for January 2023 to June 2026… 13 arXiv — NLP / Computation & Language research 17d ago Towards Reliable Medical LLMs: Benchmarking and Enhancing Confidence Estimation of Large Language Models in Medical Consultation arXiv:2601.15645v2 Announce Type: replace Abstract: Large-scale language models (LLMs) often offer clinical judgments based on incomplete information, increasing the risk of misdiagnosis. Existing studies have primarily evaluated confidence in single-turn, static settings,… 28 Vercel — AI dev-tools 17d ago How Featured's users make 100K media pitches per month on Vercel Featured on Vercel 3 engineers supporting 3 brands and 100,000+ users on Vercel Migrated 374 Sanity sites from AWS Elastic Beanstalk to Vercel AI SDK and AI Gateway power Featured's chat bot across 17 models Workflow SDK replaced custom long-running job infrastructure Featured… 38 llama.cpp releases dev-tools 17d ago b10897 ci : Update WoA CUDA 13.4 release to use 13.4.1 GA redistributables ( #28687 ) Move Windows ARM64 CUDA 13.4 builds from the Developer Preview archives to the 13.4.1 GA redistributables Website: https://llama.app Attestations:… 17 Vercel — AI dev-tools 17d ago GitHub Copilot is now available in the AI SDK harness layer The AI SDK harness layer now supports GitHub Copilot through the official @ai-sdk/harness-github-copilot adapter. The harness layer lets your application run different coding agents through the same HarnessAgent interface, so you can switch agents without changing your… 28 r/LocalLLaMA community 17d ago LoudKit: local TTS with voice cloning, 10 languages, and SDKs for Python, Swift, Go, Rust and TypeScript hey guys, I've been working on a reading app for several months now and had problems with getting good quality TTS, the options were kokoro, kitten, pocket but all of them even though they were sounding natural had some problems when listening longer. Last month I took upon… 14 r/LocalLLaMA community 18d ago Closed AI doesn't like biological research, user turns to open weight models OpenAI has decided to fully shut down a protein design project I'm working on for a client. Needless to say, open weight models are the only way forward.   submitted by   /u/Terminator857 [link]   [comments] 24 OpenAI official-blog 18d ago Introducing ChatGPT for Financial Services Introducing ChatGPT for Financial Services, combining built-in financial data and GPT-6 Astra for research, modeling, and client-ready materials. 15 arXiv — Machine Learning research 18d ago CALM: Class-wise Agreement and Label-gated Disagreement Modulation for Decentralized Federated Learning arXiv:2609.05884v1 Announce Type: new Abstract: Conventional federated learning relies on parameter averaging, which forces clients to be doubly homogeneous: all must run an identical architecture, and accuracy degrades when local data are non-IID. Decentralized federated… 38 arXiv — Machine Learning research 18d ago Machine Learning for Pre-Culture ESBL Risk Stratification to Guide Empiric Antibiotic Selection: A 12-Hospital Study of Enterobacteriaceae Cultures arXiv:2609.05970v1 Announce Type: new Abstract: Empiric antibiotic therapy for suspected ESBL-producing Enterobacteriaceae must be selected 48-72 hours before culture results, forcing clinicians to choose between undertreating resistant infections and overusing carbapenems that… 12 arXiv — Machine Learning research 18d ago Granular-Ball Quantum Clustering for Resource-Efficient and Robust Learning arXiv:2609.06016v1 Announce Type: new Abstract: Quantum clustering aims to exploit quantum feature representations to uncover complex data structures beyond conventional Euclidean geometry. Yet this sample-level kernel construction requires O(n^2) quantum circuit executions for… 20 arXiv — Machine Learning research 18d ago FedSubMuon: Communication-Efficient Federated LLM Fine-Tuning via Structured Subspace Muon arXiv:2609.06073v1 Announce Type: new Abstract: Federated fine-tuning adapts large language models (LLMs) to decentralized client data, but its scalability in cross-device training is often limited by the high communication cost. Muon is an optimizer that improves optimization… 18 arXiv — Machine Learning research 18d ago PhenoBench: Mapping What a Deeply Phenotyped Human Cohort Can Tell Us arXiv:2609.06080v1 Announce Type: new Abstract: Deeply phenotyped cohorts combine clinical, imaging, molecular, and wearable observations across timescales from seconds to years. This breadth can reveal which measurements inform which health-related questions, but heterogeneous… 25 arXiv — Machine Learning research 18d ago Rethinking One-Shot Federated Graph Learning: Training-Free Statistical Estimation arXiv:2609.06154v1 Announce Type: new Abstract: One-shot federated graph learning generally aims to train Graph Neural Networks (GNNs) across clients with disconnected subgraphs in a single communication round. Existing methods predominantly design advanced optimization… 8 arXiv — Machine Learning research 18d ago Parameterized and Streaming Algorithms for Euclidean Fair $k$-Center Clustering arXiv:2609.06384v1 Announce Type: new Abstract: Motivated by the growing importance of fairness in machine learning, fair $k$-center clustering has attracted considerable research attention as a fundamental problem. In this problem, a dataset is partitioned into $m$ disjoint… 8 arXiv — Machine Learning research 18d ago Structural Entropy-Driven Graph Diffusion Generation for One-Shot Federated Graph Learning arXiv:2609.06499v1 Announce Type: new Abstract: One-shot federated graph learning (FGL) requires the server to estimate client contributions from highly compressed information, yet conventional volume-based weighting captures the amount of client data while overlooking how its… 5 arXiv — Machine Learning research 18d ago DrugReason: Dynamic Multi-View Reasoning over Knowledge Graph and Language Evidence for Drug Repurposing arXiv:2609.06779v1 Announce Type: new Abstract: Drug repurposing aims to identify new therapeutic uses for existing compounds and, compared with de novo drug discovery, offers a faster and more cost-effective path to clinical translation. However, the space of candidate… 18 arXiv — NLP / Computation & Language research 18d ago SymbolicLight V2: Hybrid Neuromorphic Architecture and Sparse Execution for Low-Energy Language Inference arXiv:2609.09772v1 Announce Type: new Abstract: SymbolicLight V2 combines sparse event computation with continuous-state processing in a hybrid neuromorphic language architecture. Extending V1's spike-gated dual paths, it adds graded signed events at further projections and… 10 arXiv — NLP / Computation & Language research 18d ago MedDeID enables locally governed clinical-text de-identification from real or synthetic training data arXiv:2609.10049v1 Announce Type: new Abstract: Clinical notes contain personally identifiable information (PII), restricting reuse for research and medical AI, especially when data cannot leave an institution. We developed MedDeID, an on-premises framework combining in-house… 36 arXiv — NLP / Computation & Language research 18d ago AgenticGen: Reward-Guided Agentic Video Generation for Advertising arXiv:2609.09187v1 Announce Type: cross Abstract: Advertising video generation is not only a video synthesis task, but also a product-conditioned reasoning problem whose success is measured by online business metrics. Recent video foundation models can generate realistic clips… 23 arXiv — NLP / Computation & Language research 18d ago An Efficient and Effective Agentic Group Shilling Attack on Recommender Systems arXiv:2609.09551v1 Announce Type: cross Abstract: Recommender systems have become core infrastructure for modern online platforms, personalizing content at scale and strongly influencing what users see, click on, and purchase. However, this dependence on user interaction also… 19 r/MachineLearning community 18d ago I tried to make a real fly connectome learn to play Pong. It didn't — and auditing why turned out to be way more interesting than if it had worked [p] You've probably seen the fly-brain-plays-Doom / Minecraft / Beat Saber clips going around this week, from the new MaleCNS v1.0 connectome release (166k neurons, real EM reconstruction, not a toy model). Cool clips. Nobody seemed to be checking whether any of it works though,… 30 Simon Willison community 18d ago Quoting Calif Research Today, we're releasing a demo of WeWorm, the first zero-click worm to spread through WeChat calls across iOS and Android. [...] The victim does not need to answer the call, or interact with their phone at all. Even if they do answer, they hear nothing, and the exploit still… 31 r/LocalLLaMA community 18d ago Solved: LLM inference on Windows was 2–3x slower when the server window wasn't focused The fix: run the server detached/headless instead of keeping it attached to a console window. RTX 5090, ~27B NVFP4 model via ninfer: Terminal focused: 130–200 tok/s Terminal unfocused: 50–60 tok/s Click the terminal → immediately back to 130+ tok/s At first I thought GPU… 17 Hugging Face Daily Papers research 18d ago NOAH: Learning the Full Patient Journey. A Longitudinal Multimodal Time-Aware Model for Representation and Forecasting Abstract NOAH is a generative transformer that models full multimodal patient journeys with continuous time dynamics and stochastic latent states, enabling forecasting, zero-shot classification, and counterfactual simulation across diverse clinical data. Generated by… 11 r/LocalLLaMA community 18d ago SOTA ImageGen Locally NVIDIA Cosmos3(64B) INT4 quants CUDA/MLX Cosmos3 INT4 T2I + I2V on Apple Silicon — code, weights and a Grok comparison GitHub - https://github.com/gtrg55/cosmos3-quant-mlx-cuda HF weights - https://huggingface.co/JuliaML/Cosmos3-Super-Text2Image-4Step-INT4-G64-BF16 Single clip took approximately 5m on M4 MAX 128 GB Mac… 26 Hugging Face Daily Papers research 18d ago SynthGait-19K: A Physically Grounded Synthetic Video Dataset for Gait Parameter Estimation Abstract SynthGait-19k is a large synthetic video dataset for gait analysis that enables benchmarking of video-based gait estimation and shows synthetic supervision transfers to real data. Generated by thinkingmachines/Inkling-Small Accurate estimation of clinically meaningful… 9 Hugging Face Daily Papers research 19d ago Harnessing CLIP and DINO: An Uncertainty-Aware Cascaded Fusion Network for Generalizable Deepfake Image Detection Abstract UCF-Net improves deepfake detection by fusing CLIP and DINO representations with uncertainty-weighted hierarchical feature aggregation, achieving stronger cross-domain generalization. Generated by thinkingmachines/Inkling-Small The growing realism and accessibility of… 31 Hugging Face Daily Papers research 19d ago Steering Geometry: Validating Human Value Geometry in LLM Steering Space Abstract Activation steering vectors in large language models encode theory-aligned human value geometry when derived via distribution-driven methods, with geometric fidelity scaling with model size but declining after instruction tuning. Generated by… 7 Vercel — AI dev-tools 19d ago v0 adds one-click integrations for email, auth, search, and databases We're working toward bringing parity across Vercel integration and v0 , starting with Resend , Amazon OpenSearch , MongoDB Atlas , Algolia and Clerk . Prompt v0 with what you want to build, and when your prompt requires a provider, v0 renders a connect card in the chat. With… 8 Vercel — AI dev-tools 19d ago You can now read and search changelogs from the CLI You and your agents can now read and search the Vercel changelog feed from your terminal using vercel changelog . Coding agents can use this command to discover new Vercel products and features, find updates relevant to your project, and read full announcements to inform their… 16 Vercel — AI dev-tools 19d ago Vercel Sandbox routing is now 18x faster globally Requests to Vercel Sandbox public domains are now routed 18x faster. Domains created with the sandbox.domain() SDK call are now resolved from the nearest regional replica, instead of a single centralized store. Incoming requests reach the process running in the sandbox with less… 17 The Information — AI news-outlet 19d ago White House Declines Senate Democrats’ Request for AI Policy Clarity Attempts by Senate Democrats to get more clarity about the White House’s AI policies have gone nowhere. A White House official responded last Friday to an August letter from a group of Senate Democrats asking for more details about the White House’s voluntary AI testing… 21 Hacker News — AI on Front Page community 20d ago I've operated petabyte-scale ClickHouse clusters for 5 years Article URL: https://www.tinybird.co/blog/what-i-learned-operating-clickhouse Comments URL: https://news.ycombinator.com/item?id=49601138 Points: 214 # Comments: 78 14 Hugging Face Daily Papers research 21d ago τ^τ-Bench: An Environment for End-To-End, Realistic Agent Construction Abstract The τ^τ-bench benchmark evaluates coding agents on building real-world customer-service agents from business records, client requirements, and production APIs, revealing substantial gaps versus expert performance. Generated by thinkingmachines/Inkling-Small LLM agents… 10 arXiv — Machine Learning research 21d ago Nested Inductive Bias Framework for SPD Manifold Learning arXiv:2609.04466v1 Announce Type: new Abstract: In Geometric Deep Learning, inductive biases serve two primary functions: enforcing manifold constraints and embedding relational priors. Currently, representation learning on SPD manifolds frequently relies on pullback Euclidean… 7 arXiv — Machine Learning research 21d ago Mitra-v2 Technical Report arXiv:2609.04540v1 Announce Type: new Abstract: We introduce Mitra-v2, a tabular foundation model that delivers state-of-the-art performance on real-world classification and regression problems, from credit-risk scoring and clinical prediction to equipment-failure detection and… 9 arXiv — Machine Learning research 21d ago Resilience Beyond Stationary Client Unavailability: Unlocking Efficient and Unbiased Federated Learning arXiv:2609.04763v1 Announce Type: new Abstract: Due to resource constraints or external and internal uncertainties, clients in real-world federated learning systems are often intermittently available edge devices. In highly dynamic environments, the parameter server lacks prior… 37 arXiv — Machine Learning research 21d ago PACE: Propagation-Aware Collaborative Correction for One-Shot Personalized Federated Graph Learning arXiv:2609.04832v1 Announce Type: new Abstract: Client heterogeneity creates both an opportunity and a risk in personalized federated graph learning. Knowledge held by other subgraphs may complement a receiver's Local model, but an incompatible transfer can override reliable… 18 arXiv — Machine Learning research 21d ago Client-Side Probing of Deleted Ridge Statistics in Federated Unlearning arXiv:2609.04475v1 Announce Type: cross Abstract: Federated unlearning aims to remove a client's data from a shared model without retraining from scratch. Some efficient systems make deletion exact by storing compact, additive summaries of the training features and broadcasting… 8 arXiv — NLP / Computation & Language research 21d ago VERGE: Verification-Enhanced Refinement for Grounded Extraction of Early-Onset Colorectal Cancer Symptoms in Clinical Notes arXiv:2609.04366v1 Announce Type: new Abstract: Early-onset colorectal cancer is increasing among younger adults, yet red-flag symptoms in this age group have no evidence-based guidelines for follow-up testing, and structured encounter data do not capture the detail needed to… 17 arXiv — NLP / Computation & Language research 21d ago PetQA: Benchmarking Veterinary Knowledge and Clinical Reasoning arXiv:2609.04598v1 Announce Type: new Abstract: We introduce PetQA, a Korean long-form question-answering (QA) benchmark for evaluating veterinary knowledge and clinical reasoning in large language models (LLMs) and large vision-language models (LVLMs). PetQA contains 10,076… 5 Page 4 of 10 · 500 articles ← Newer Older →