News / #developer-tool Tag Developer Tool 500 articles archived under #developer-tool · RSS Sign in to follow arXiv — Machine Learning research 6d ago A Comparative Framework for Evaluating Foundation Models on Tabular Data: A Case Study in Healthcare arXiv:2609.22154v1 Announce Type: new Abstract: Tabular data is the most common format in clinical practice, encompassing laboratory results, medication records, diagnostic codes, and patient demographics. As foundation models for tabular data have grown in number and variety, a… 4 arXiv — Machine Learning research 6d ago From Latent Biomarkers to Clinical Rules: Embedding-Guided Rule Mining and Attribution-Based Translation for Interpretable Tabular Learning arXiv:2609.22155v1 Announce Type: new Abstract: Clinical decision support tools are most useful when accurate predictions are accompanied by understandable explanations. Rule-based models provide transparency, but rules derived directly from raw clinical measurements may miss… 12 arXiv — Machine Learning research 6d ago The Effect of Quantization on Clinical Benchmarks: Accuracy and Safety Across Model Families arXiv:2609.22216v1 Announce Type: new Abstract: Quantization enables deployment of large language models on resource-constrained clinical edge devices, but its effect on clinical accuracy and safety remains understudied. We evaluate five 7-8B parameter models at FP16, GPTQ-INT8,… 37 arXiv — Machine Learning research 6d ago Joint Domain-Class Modeling for Federated Learning Under Feature Skew arXiv:2609.22932v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training without centralizing private data, but performance often degrades under feature skew: clients share labels while the conditional input distributions $p_i(x\!\mid\!y)$… 26 arXiv — NLP / Computation & Language research 6d ago AI-inferred expressed well-being and collective-action discourse in climate-change campaigns on X arXiv:2609.22096v1 Announce Type: new Abstract: Climate campaigns are often evaluated through attention and mobilization, but less is known about the well-being language that accompanies them. Whether campaign periods alter positive affect and hope, and whether happiness aligns… 35 arXiv — NLP / Computation & Language research 6d ago A Multi-Agent Pipeline for Source-Grounded Synthetic Note Generation from Longitudinal Structured EHR arXiv:2609.22164v1 Announce Type: new Abstract: Structured EHR is abundant but sparse, coded, and difficult to use directly for note-centric clinical modeling. We present MedNotes, a multi-agent synthetic data generation pipeline that converts longitudinal structured EHR into… 34 arXiv — NLP / Computation & Language research 6d ago Knowledge Graph-Augmented Ambient AI for Clinical Note Generation arXiv:2609.22239v1 Announce Type: new Abstract: Ambient AI is increasingly adopted in healthcare to automatically generate clinical notes from patient-clinician conversations, with the potential to substantially reduce clinician documentation burden. However, generated notes may… 19 arXiv — NLP / Computation & Language research 6d ago Analyzing Public Discourse on Urbanism: Topic Clustering, Sentiment Analysis and Retrieval-Augmented Generation using YouTube Comments arXiv:2609.22705v1 Announce Type: new Abstract: Online discourse about urban issues - walkability, cycling infrastructure, public transit, housing density, and street safety - is voluminous but unstructured, and existing city-evaluation tools capture none of it. We present a… 24 arXiv — NLP / Computation & Language research 6d ago Clinical Domain Classification from Medical Transcriptions arXiv:2609.22734v1 Announce Type: new Abstract: Clinical domain classification plays an important role in organizing and analyzing large volumes of unstructured medical text. However, medical transcription datasets are often highly imbalanced, which can substantially degrade… 33 arXiv — NLP / Computation & Language research 6d ago LLMs Anchor on Chief Complaint and Fail to Integrate Evidence in Sequential Clinical Triage arXiv:2609.22904v1 Announce Type: new Abstract: Triage in the emergency department (ED) is a sequential decision process that unfolds turn by turn. Existing evaluations of large language models (LLMs) for triage use completed retrospective records and report performance close to… 13 Ars Technica — AI news-outlet 6d ago Muse, Meta's extraordinarily privileged AI assistant, has a serious 0-day A simple ClickFix attack is only one way to completely hijack the new agent. 16 TechCrunch — AI news-outlet 6d ago With Tabby, a former accountant is using AI to make accountants obsolete Tabby is designed to be a real-time bookkeeping interface, handling clients’ paperwork as it gives them up-to-the-minute data on their business’s profit and loss. 30 Hacker News — AI on Front Page community 6d ago Fable 5 – Median thinking declined in August Article URL: https://twitter.com/Lon/status/2101793422487204027 Comments URL: https://news.ycombinator.com/item?id=49789224 Points: 244 # Comments: 163 24 r/LocalLLaMA community 6d ago [Splash Engine] Qwen3.8-27B in native 8-bit at 37–55 tok/s on Apple Silicon: Extending Splash to Q8, 256k context scaling, and the "Reasoning Cliff" https://preview.redd.it/nulsv53o8vqh1.png?width=4500&format=png&auto=webp&s=74765dbd409f4c221640f9f6000a685f6fdbb242 Spent weekend benchmarking the Splash engine (by Incoai) and extending its architecture to native 8-bit on Apple Silicon (M5 Pro, 64 GB unified memory). Splash is… 10 llama.cpp releases dev-tools 7d ago b11068 metal : fix deprecation warnings from macOS 27 SDK ( #29136 ) Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/48890258 macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS… 29 arXiv — Machine Learning research 7d ago FedeRage: Provably Convergent Agnostic Federated Learning under General Client Drift arXiv:2609.21057v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training without sharing raw data, but its performance degrades under non-IID data and stochastic client participation. Remedies built on classical Federated Averaging (FedAvg)… 29 arXiv — Machine Learning research 7d ago M2G-LLM: Enhancing Clinical Prediction via Multimodal Graph Reasoning and LLM Context Injection arXiv:2609.21164v1 Announce Type: new Abstract: Integrating diverse data modalities --- such as clinical notes, laboratory results, and medical imaging --- is essential for advancing clinical decision-making. While Large Language Models (LLMs) have shown remarkable performance… 25 arXiv — Machine Learning research 7d ago Riemannian Neural Hamiltonian Flows: Geodesic Symplectic Transport and Interpretability arXiv:2609.21647v1 Announce Type: new Abstract: Hamiltonian normalizing flows are attractive generative models because their phase-space maps are invertible and volume preserving, but most neural constructions are formulated in Euclidean space. We introduce Riemannian Neural… 17 arXiv — Machine Learning research 7d ago Intervention Granularity Matters: Coherent Treatment Bundles in Counterfactual Simulation with Clinical World Models arXiv:2609.21906v1 Announce Type: new Abstract: Counterfactual simulation with a clinical world model means fixing a patient's history, changing the treatment, and reading off the predicted response. Doing so requires deciding what counts as one intervention. In clinical… 8 arXiv — NLP / Computation & Language research 7d ago HERMES: Contrast-Aware Knowledge Graph Reasoning from Clinical Notes for Patient Outcome Prediction arXiv:2609.20825v1 Announce Type: new Abstract: Clinical predictive models often rely on structured Electronic Health Record data, such as time-series and procedure codes. While recent approaches have begun leveraging unstructured clinical notes, they typically encode them as… 20 arXiv — NLP / Computation & Language research 7d ago PhysioBench: A Unified Benchmark for Physiological Signal Question Answering arXiv:2609.20836v1 Announce Type: new Abstract: Physiological signals support diverse clinical and monitoring tasks, yet existing physiological signal foundation models typically require task-specific adaptation for each task. Natural language provides a common interface for… 34 arXiv — Machine Learning research 7d ago Reconstruction of 4D Mitral Regurgitation Hemodynamics from Sparse Planar Data using Deep Operator Networks with Test-Time Adaptation arXiv:2609.20857v1 Announce Type: cross Abstract: Quantifying mitral regurgitation severity remains limited by the assumptions of clinical flow convergence methods, while high-fidelity simulation and volumetric velocimetry are too slow for routine use. We investigate whether a… 26 arXiv — NLP / Computation & Language research 7d ago From Discharge Notes to Patient Understanding: Persona-Grounded, Open-Ended Simulation of LLMs as Discharge Educators arXiv:2609.20827v1 Announce Type: new Abstract: Hospital discharge education is an interactive teaching task: a clinician adapts a discharge plan to a patient's literacy, recall, and personality. Existing LLM evaluations target static or artifact-generation tasks and do not… 17 arXiv — NLP / Computation & Language research 7d ago Reading Anxiety or Reading the Label? Comparing Fine-Tuned and Frontier Models for Anxiety Detection on Social Media arXiv:2609.20847v1 Announce Type: new Abstract: Anxiety is among the most common mental health conditions, and people often write about it online well before seeking clinical help. Practitioners building detection tools face a concrete choice: call a frontier commercial model,… 22 arXiv — NLP / Computation & Language research 7d ago Consistent Relexicalization of Clinical Documents using Graph-Based Approach arXiv:2609.21387v1 Announce Type: new Abstract: Relexicalization is a pivotal technique in clinical NLP, as it facilitates robust masking of sensitive information while synthesizing datasets that retain high-fidelity, real-world characteristics. However, preserving structural… 28 arXiv — NLP / Computation & Language research 7d ago TrialAtlas: Multi-Agent Research Organization for Clinical Trial Design and Optimization arXiv:2609.21859v1 Announce Type: new Abstract: Nearly 90% of drugs entering clinical development ultimately fail, despite billions of dollars in investment. Pharmaceutical companies therefore rely on clinical development planning (CDP) and probability of technical and… 19 arXiv — NLP / Computation & Language research 7d ago Clinician-Grounded Quality Assurance for AI-Assisted Psychiatric Intake arXiv:2609.21149v1 Announce Type: cross Abstract: Before patients can use AI-assisted psychiatric intake systems, health systems need practical ways to routinely evaluate these tools against their clinical standards for quality assurance. Because clinicians may use different… 15 arXiv — NLP / Computation & Language research 7d ago Hierarchical attention interpretation: an interpretable speech-level transformer for bi-modal depression detection arXiv:2309.13476v3 Announce Type: replace Abstract: Depression is a common mental disorder. Automatic depression detection tools using speech, enabled by machine learning, help early screening of depression. This paper addresses two limitations that may hinder the clinical… 18 r/LocalLLaMA community 7d ago ZCode is now open source ZCode is now open source , and the reported security issues have been addressed. Source code: https://github.com/zai-org/ZCode The repo includes its desktop app, web workspace, backend, Agent CLI, and runtime. Official announcement: In response to the ZCode product security… 24 Vercel — AI dev-tools 7d ago AI Gateway now supports TypeSafe clients and an HTTP API for Jev You can now call Jev from TypeSafe AI through AI Gateway using an existing TypeSafe client or the HTTP API, in addition to the AI SDK. TypeSafe client: Point an existing TypeSafe client at AI Gateway without changing its evaluation calls. HTTP API: Call Jev directly from any… 10 The Information — AI news-outlet 7d ago Why Apple Is Betting on Foldable iPhones When new Apple CEO John Ternus showed off the first version of a foldable iPhone at an event earlier this month, there was a good chance many Americans watching clips of the video stream had never thought hard about buying such a device. But Apple is betting that the same demand… 13 r/LocalLLaMA community 9d ago Steer LLMs and Agents at the Token Level: An interactive tool for token visualization & control, model inspection and data annotation. onPanda is designed for geeks, power users, curious minds, and engineers. Its UI is built for deep exploration and efficient data annotation. - The core loop is simple: hover over a token → click an alternative or edit freely → continue generation. You can edit every part of… 21 arXiv — Machine Learning research 10d ago Personalized Federated Hierarchical Gaussian Processes for Privacy-Preserving Modeling of Heterogeneous Distributed Systems arXiv:2609.19337v1 Announce Type: new Abstract: We present Personalized Federated Hierarchical Gaussian Processes (pFedHGP) for probabilistic regression and classification when data are distributed across heterogeneous clients. Each client's latent function decomposes into (i) a… 7 arXiv — Machine Learning research 10d ago Stiefel Attention: When the Geometry of Transformer Projection Matrices Dominates Optimizer Choice---and When It Does Not arXiv:2609.19363v1 Announce Type: new Abstract: The query and key projections $\WQ,\WK$ in attention are almost always trained by Euclidean optimizers with no constraint on their geometry. We constrain them to the Stiefel manifold and optimize them there with a Riemannian Adam… 29 arXiv — Machine Learning research 10d ago Machine-Learning Assessment of the Predictive Value of Inflammatory Biomarkers for Cognitive Impairment in an Older Hispanic Adult Cohort arXiv:2609.19374v1 Announce Type: new Abstract: Small clinical tabular datasets require interpretable machine learning because deep learning is often impractical and ensemble models can be difficult to inspect. A key pitfall is that statistical significance does not necessarily… 38 arXiv — Machine Learning research 10d ago FedFIbOS: Fisher Importance based Optimal Submodelling for Heterogeneous Federated Learning arXiv:2609.19559v1 Announce Type: new Abstract: Heterogeneous federated learning requires clients with diverse computational capacities to collaboratively train a global model, where each client trains a capacity-constrained submodel. Existing methods select submodel parameters… 15 arXiv — Machine Learning research 10d ago Pretrained Medical Representations for the Practical Screening of Drug Repositioning Candidates arXiv:2609.19865v1 Announce Type: new Abstract: Representation learning from medical code sequences in electronic health records and medical claims data has been successful in various clinical applications, such as those regarding disease prediction. However, significant… 24 arXiv — Machine Learning research 10d ago Distributionally Robust Federated Learning with Multi-Source Data arXiv:2609.20501v1 Announce Type: new Abstract: Federated learning trains a shared model from private client data. In practice, data-generating distributions may differ, and the true mixture across clients is often unknown, making the underlying group distribution difficult to… 36 arXiv — NLP / Computation & Language research 10d ago Modality Discrepancy Transformer for Ambivalence and Hesitancy Recognition arXiv:2609.19148v1 Announce Type: new Abstract: Ambivalence and hesitancy (A/H) are affective states in which individuals express contradictory signals across facial, vocal, and linguistic channels. Automatically recognising A/H in clinical videos requires detecting cross-modal… 5 arXiv — NLP / Computation & Language research 10d ago CliniCIRCA: A Modular LLM Framework for Constructing Longitudinal Mental Health Patient Journeys from Raw EHR Narratives arXiv:2609.19585v1 Announce Type: new Abstract: In mental health care, reasoning over patient journeys is a key task for clinicians. Yet these journeys, encompassing a longitudinal progression of biological, psychological, and social events, are often spread across disparate… 37 arXiv — NLP / Computation & Language research 10d ago Harm Laundering in GPT Models: Evidence That Gender Discrimination Is Transformed Rather Than Reduced Across Safety-Trained Generations arXiv:2609.20779v1 Announce Type: new Abstract: Safety evaluations for large language models rely on surface-form classifiers that report declining harm scores across model generations. We provide evidence that this methodology is systematically incomplete: explicit… 7 arXiv — NLP / Computation & Language research 10d ago BurnRiSc: Toward Non-Invasive Burnout Screening in Open Source from Public Repository Signals arXiv:2609.19422v1 Announce Type: cross Abstract: Burnout is a chronic occupational syndrome, and open source is close to a worst case for it: maintainers absorb unbounded demand with no manager to reallocate work and no organization to notice decline. The cost is not only… 7 Vercel — AI dev-tools 10d ago Sub-second artifact deployments are now supported in Vercel CLI You and your agents can now deploy static artifacts to Vercel in under one second through Vercel CLI. Run vercel deploy to share a prototype, publish an HTML report, or preview a page created by your coding agent. Vercel automatically detects eligible deployments, and valid… 20 Vercel — AI dev-tools 10d ago The skills CLI now supports Notion hosted skills [email protected] adds Notion skills databases as an install source for agent skills . Notion skills are reusable agent skills written as Notion pages. Teams author, review, and update them in the workspace they already use, then install them into any agent the skills CLI supports.… 27 The Information — AI news-outlet 10d ago Same Flaw Found in Claude Code, Codex, Gemini CLI and GitHub Copilot While the AI industry ties itself in knots over models it thinks will pose security threats to the internet and perhaps to humanity itself, researchers are warning about flaws in AI coding agents from Anthropic, OpenAI, Google and Microsoft that pose immediate risk to… 8 r/LocalLLaMA community 11d ago How do you create and use your agentic workflow? What I want to know is how your daily interaction routine looks like. Do you wake up, take your coffee/tea/drinkname.txt, have a nice morning, then click a button and somethings starts and creates something for you? I can understand coders or data analysts for example, where AI… 16 arXiv — Machine Learning research 11d ago REVERSAL-BENCH: A Reversibility Axis and Reset Oracle for Measuring the Reset-Free RL Cliff arXiv:2609.17745v1 Announce Type: new Abstract: A central goal of autonomous reinforcement learning is continuous policy training without external resets. However, existing paradigms largely depend on underlying environmental reversibility, a property absent in real world… 38 arXiv — Machine Learning research 11d ago NeuroECG: ECGFounder-Based Deep ECG Representation for EEG-Free Neurological Prognostication After Cardiac Arrest arXiv:2609.18891v1 Announce Type: new Abstract: Neurological prognostication after cardiac arrest commonly relies on electroencephalography (EEG). However, EEG demands high clinical resources. Bedside electrocardiography (ECG) is standard and low-cost. Yet, its value for… 31 arXiv — NLP / Computation & Language research 11d ago Enhancing Extubation Failure Prediction with LLM-Derived Features from Respiratory Therapy Clinical Notes arXiv:2609.17532v1 Announce Type: new Abstract: Invasive mechanical ventilation is a lifesaving therapy, but timely, safe discontinuation is essential to preventing extubation failure (EF) and related risks to health. We present a novel approach to EF prediction that leverages… 30 arXiv — NLP / Computation & Language research 11d ago Large Language Models Versus Physicians in Traditional Chinese Medicine: A Real-World Clinical Case Evaluation arXiv:2609.17544v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being explored for clinical applications, yet their assessment for real-world traditional Chinese medicine (TCM) practice remains limited We constructed a clinical case library… 33 Page 2 of 10 · 500 articles ← Newer Older →