News / #ide Tag Ide 151 articles archived under #ide · RSS Sign in to follow Simon Willison community 9h ago sqlite-utils 4.2.1 Release: sqlite-utils 4.2.1 Fixes a crashing bug in sqlite-utils 4.2 . I'd introduced code that looks like this: from typing_extensions import Self It turned out the typing-extensions package was not listed as a dependency for sqlite-utils - it was installed by one of the other… 17 arXiv — NLP / Computation & Language research 1d ago On Weak Bisimilarities in CCSK arXiv:2608.11531v1 Announce Type: new Abstract: In the context of CCSK, a reversible extension of CCS, we study different notions of bisimilarity (strong/weak, forward-only/reversible) and highlight their differences and commonalities. In particular, for the weak reversible… 20 arXiv — Machine Learning research 2d ago Hierarchical Empirical-Bayes Naive Bayes: Minimax Smoothing and Calibration with AODE Extension arXiv:2608.11162v1 Announce Type: new Abstract: The Naive Bayes (NB) classifier remains a standard choice for categorical data, yet its widely used smoothing rules, such as Laplace, Lidstone, Krichevsky-Trofimov, and the $m$-estimate, all prescribe a fixed smoothing strength… 14 arXiv — NLP / Computation & Language research 2d ago Cracks in the Foundation: Seemingly Minor Architectural Choices Impact Long Context Extension arXiv:2608.10296v1 Announce Type: new Abstract: One might imagine that architectural variations within the dense transformer paradigm have a limited effect on accuracy. However, we demonstrate that this is not the case in the long context setting. Specifically, we show that a… 32 Hugging Face Daily Papers research 2d ago TSDS-Toolbox: A Toolbox for Measuring Time-Series Dataset Similarity Abstract A unified toolbox enables reproducible comparison and extension of time-series dataset similarity methods for forecasting, classification, and generation tasks. Generated by thinkingmachines/Inkling-Small The rapid advancement of artificial intelligence (AI) has… 10 TechCrunch — AI news-outlet 2d ago Spotify will label ‘AI Persona’ profiles and exclude their music from recommendations Spotify is introducing “AI Persona” labels for artist profiles that represent AI-generated identities and will exclude their music from editorial, algorithmic, and personalized recommendations by default. 24 vLLM releases dev-tools 3d ago v0.27.1: [CI] Limit Arctic import check to x86 test images The arm64 test lockfile intentionally omits arctic-inference, so only validate its native extension on platforms where the package is installed. Co-authored-by: OpenAI Codex [email protected] Signed-off-by: khluu [email protected] 19 arXiv — NLP / Computation & Language research 3d ago Mawqif-v2: An Arabic Benchmark Dataset for Cross-Target Stance Detection arXiv:2608.09539v1 Announce Type: new Abstract: Publicly available Arabic datasets for target-specific stance detection remain limited, particularly for evaluating cross-target generalization. This paper presents the Mawqif-v2 Extension, consisting of 996 manually annotated… 9 arXiv — NLP / Computation & Language research 3d ago REFRAMED: Towards Realistic Audio Description Generation for Movies arXiv:2608.09765v1 Announce Type: new Abstract: Audio Description (AD) is a verbal narration of key visual content in videos, enabling access for visually impaired audiences. Unlike standard video captioning, AD is a structured editorial task: descriptions must be inserted into… 35 r/LocalLLaMA community 6d ago Anyone else amped up over Qwen 3.8? I’ve been using 3.6 27B Q4, and that quant is fast on an M5. The code has been average, but consistently “good enough.” And, after a year, I can see home LLMs being served at home much like streaming music was introduced. A simple browser extension and all your queries go… 14 arXiv — NLP / Computation & Language research 7d ago Building Open-Retrieval Conversational Question Answering Systems by Generating Synthetic Data and Decontextualizing User Questions arXiv:2507.04884v2 Announce Type: replace Abstract: We consider open-retrieval conversational question answering (OR-CONVQA), an extension of question answering where system responses need to be (i) aware of dialog history and (ii) grounded in documents (or document fragments)… 4 llama.cpp releases dev-tools 8d ago b10293 ci : onboard AMD ROCm CI with gfx1151 fixes ( #26544 ) ci: prepare for amd rocm ci Signed-off-by: Aaron Teo [email protected] ci: fix editorconfig-checker Signed-off-by: Aaron Teo [email protected] ci: fix device not recognised Signed-off-by: Aaron Teo [email protected] ci:… 23 arXiv — NLP / Computation & Language research 9d ago Beyond Initialization Loss: A Systematic Study of Token Embedding Initialization Strategies for LLM Vocabulary Extension arXiv:2608.03494v1 Announce Type: new Abstract: Vocabulary extension is an efficient way to adapt pretrained large language models (LLMs) to new languages, but the initialization of newly added token embeddings can strongly affect continued pre-training (CPT) efficiency. We… 11 arXiv — NLP / Computation & Language research 9d ago LACE: Large Language Model Aided Multi-Agent Framework for Agile RISC-V Instruction Extension arXiv:2608.02915v1 Announce Type: cross Abstract: Domain-specific Instruction Set Architecture eXtensions (ISAX) are widely adopted in the RISC-V ecosystem to accelerate emerging workloads, but implementing and validating ISAXes across different cores remains slow and… 32 arXiv — Machine Learning research 10d ago Stabilized Best-of-$K$ Training for Neural Combinatorial Optimization arXiv:2608.00296v1 Announce Type: new Abstract: Leader Reward modifies POMO training to emphasize the best trajectory produced by repeated inference. We test a narrow extension: replace its binary leader/non-leader distinction with a stabilized rank signal indexed by a sampling… 26 Vercel — AI dev-tools 10d ago Give your eve agent a browser Your eve agent can now navigate the web like a human with agent-browser . The @agent-browser/eve extension gives any eve agent a full set of browser tools: navigate pages, read content, click, fill forms, take screenshots, and inspect console and network activity. Everything… 33 Simon Willison community 13d ago Slack Emoji Maker Tool: Slack Emoji Maker I wanted to create a new Slack emoji, and their tool recommends a square that's 128x128 and has a transparent background... so I had Fable build me this simple image editor against those requirements. Tags: tools , slack 19 llama.cpp releases dev-tools 13d ago b10201 ggml-webgpu: improve flash_attn_vec for quantized KV at long contexts ( #25956 ) improve fa of quantized kv cache Fix some bugs and some comments. fix v type check and some comments Fix build error caused by rebasing editorconfig checking pass Website: https://llama.app… 13 Hugging Face Daily Papers research 14d ago Pedestrian Archetypes Extension -- More Pedestrian Models for Autonomous Vehicle Safety Testing Abstract In our prior work, Pedestrian Archetypes, we defined pedestrian archetypes as collections of behaviors that uniquely identify a specific type of pedestrian. The first paper proposed 12 pedestrian archetypes, including the Wanderer, Drunk, Distracted, Flash, Indecisive,… 6 arXiv — NLP / Computation & Language research 14d ago CACHE-UK: A Stability-Aware Memory Editor for Sequentially Updated Quantized LLMs in Finance arXiv:2607.28292v1 Announce Type: new Abstract: Large Language Models (LLMs) deployed in dynamic financial environments face a critical challenge: maintaining factual accuracy as market conditions, regulations, and corporate facts change continuously. While 4-bit quantization… 35 r/LocalLLaMA community 15d ago A slide deck you can edit with a local model or in Chrome — the whole deck is a JSON block in one HTML file (~640KB with editor and viewer included) Over the past few months, our team has been building more and more slidedecks using web frontend technologies with coding harnesses, but a common complaint is to make even small edits we need to edit the code either manually or via the harness. To avoid this loop, I ended up… 17 arXiv — NLP / Computation & Language research 16d ago Research Report on Noise-Shaped One-Bit Coefficients in Discrete Polynomial Fourier Extension arXiv:2607.24868v1 Announce Type: new Abstract: This report studies noise-shaped one-bit coefficients in normalized discrete polynomial Fourier extension. For first-order Sigma-Delta quantization, the error is written as $e_k=u_k-q_k=\Delta v_k$ with a uniformly bounded state.… 25 r/MachineLearning community 17d ago Pattern Recognition (Elsevier): "With Editor" status date changed, but status didn't. Is this normal? [R] Hi everyone, I have a manuscript under review at Pattern Recognition (Elsevier) , and I'm a bit confused about the Editorial Manager status. My timeline is: Submitted: May 26, 2026 Re-submitted after making corrections, July 1: Status "With Editor" July 22: The Status Date… 35 Hugging Face Daily Papers research 17d ago Oxygen-TryOn: Fashion-Native Foundation Model for Any-item Virtual Try-On Abstract We present Oxygen-TryOn, a unified foundation model for any-item virtual try-on. Rather than repurposing a general-purpose image editor, Oxygen-TryOn is fashion-native, built for try-on through a dedicated data engine and try-on-specific training. Given one or more… 6 r/LocalLLaMA community 17d ago NYT: Protect America’s lead in the A.I. race. “China is working hard to catch up, and the United States should take steps to keep its advantage. Most important, it should continue to prohibit American companies from selling the most advanced chips and equipment to China.” — The editorial board Ridiculous piece from the… 27 arXiv — NLP / Computation & Language research 21d ago Demonstrating GenDB: Instance-Optimized and Customized Query Processing Code Generation via LLM Agents arXiv:2607.20630v1 Announce Type: cross Abstract: Traditional query processing engines require continuous development and extensions to support new techniques and user requirements, and in some cases, entirely new systems must be built from scratch. However, these engines are… 34 Hugging Face Daily Papers research 22d ago Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning Abstract Reinforcement learning with verifiable rewards (RLVR) has substantially improved language-model reasoning, yet its extension to vision-language models remains constrained by the lack of training data that are simultaneously broad, exactly verifiable, and reproducible.… 36 arXiv — Machine Learning research 22d ago Interpretable Fuzzy Rule-Based Regression Extension for Ex-Fuzzy Library arXiv:2607.20277v1 Announce Type: new Abstract: Machine learning models achieve high predictive accuracy in regression tasks, but their deployment in safety-critical and regulated domains requires interpretability. While fuzzy rule-based systems offer transparent, linguistically… 23 arXiv — NLP / Computation & Language research 22d ago The Impact of Editorial Intervention on Detecting Native Language Traces arXiv:2605.10216v2 Announce Type: replace Abstract: Native Language Identification (NLI) is the task of determining an author's native language (L1) from their non-native writing. With the advent of human-AI co-authorship, learner texts are routinely corrected and rewritten by… 19 Vercel — AI dev-tools 22d ago GitHub tools are now an installable eve extension You can now add GitHub tools to your eve agent as an extension . Add the package, drop one file in agent/extensions/ , and your agent gets all 42 tools with Vercel Connect auth, presets, and approval rules built in. Install @github-tools/eve-extension : Then register it from a… 25 arXiv — Machine Learning research 23d ago A Self-Evolving Default Action for Cooperative Tasks with Continuous Action Space arXiv:2607.18597v1 Announce Type: new Abstract: Counterfactual credit assignment has proven effective in multi-agent reinforcement learning (MARL) for discrete action spaces, yet its extension to continuous-action cooperative tasks remains challenging. Existing methods that… 32 arXiv — NLP / Computation & Language research 23d ago A Situational Speech Synthesizer for Yoruba: System Design, Phonological Rule Architecture, and Orthographic Extensions for Contour arXiv:2607.18317v1 Announce Type: cross Abstract: We present TTSYoruba, a rule-based concatenative diphone speech synthesizer for Yoruba, deployed at online as part of the YorubaName.com open dictionary of Yoruba personal names. The system takes tone-marked Yoruba text as input… 32 Vercel — AI dev-tools 23d ago Extend eve agents with installable extensions You can now package tools, connections, skills, instructions, and hooks into extensions that any eve agent can import. Extensions can be published to package registries like npm, then installed, versioned, and upgraded like any other project dependency. A browser-use extension… 11 r/LocalLLaMA community 23d ago pi 0.81.0 adds support for llama.cpp pi 0.81.0 now has integrated support for llama.cpp (llama-server router). https://pi.dev/docs/latest/llama-cpp This seems to be able to replace the huggingface/pi-llama extension and/or manually managing models in the config.   submitted by   /u/popoppypoppylovelove… 19 arXiv — Machine Learning research 24d ago Periodic Bootstrap Thompson Sampling For Periodically Non-Stationary Bandit Problems arXiv:2607.16986v1 Announce Type: new Abstract: This paper introduces Periodic Bootstrap Thompson Sampling (PBTS), an innovative extension of the classic Thompson Sampling (TS) algorithm tailored for bandit problems with periodic non-stationarity. Conventional TS accumulates all… 38 r/LocalLLaMA community 25d ago How did I do boys? Got it for 88k PHP (~$1,443). Owner was a video editor who recently upgraded to a Macbook Pro M5 48GB SPECIFICATIONS: AMD Ryzen 9 5950X (16 Cores / 32 Threads) ASUS ROG Strix X570-I Gaming (Mini-ITX) RTX 3090 24GB 64GB DDR4 RAM Kingston NV2 2TB NVMe SSD Thermaltake SFX-L 1000W… 25 arXiv — Machine Learning research 25d ago qZACH-ViT: Quantization-Aware Intrinsic Explanations with Recursive Attribution-Stabilized Optimization arXiv:2607.15421v1 Announce Type: new Abstract: Compact medical-image classifiers need efficiency and interpretable evidence, yet these goals are often addressed separately. We introduce qZACH-ViT, a quantization-aware extension of the zero-token (CLS-token-free), position-free… 19 arXiv — Machine Learning research 25d ago AV-JEPA: Extending LeJEPA to Audio-Visual Self-Supervised Learning arXiv:2607.15295v1 Announce Type: cross Abstract: We present AV-JEPA, an elegant multimodal extension of LeJEPA to audio-visual self-supervised learning. Using an early-fusion Vision Transformer and modality dropout as masking, the model is trained to align the embeddings of… 31 arXiv — Machine Learning research 28d ago Counterfactuals for Feature-Weighted Clustering arXiv:2607.14719v1 Announce Type: new Abstract: Counterfactual explanations provide local, interpretable insight by identifying changes to an input that would alter its assigned outcome. Although well established in supervised learning, their extension to clustering is less… 12 Smol AI News news-outlet 1mo ago not much happened today **OpenAI's agent products** saw a **2.5x weekly usage growth** driven by **Codex + ChatGPT Work** and demand for **GPT-5.6 Sol**. JetBrains adopted Codex as a recommended agent, while LangChain enhanced tracing and observability across multiple tools. **PrismML released Bonsai… 36 arXiv — NLP / Computation & Language research 1mo ago A Stepwise Questioning Expert-Editor Multi-Agent Framework for Long-Document Summarization arXiv:2607.10390v1 Announce Type: new Abstract: Although large language models (LLMs) have shown promising potential in news summarization tasks, their performance on long-document summarization remains challenging as their length often exceeds the input limits. As the agent… 14 arXiv — NLP / Computation & Language research 1mo ago The Nuts and Bolts of Natural Language to SQL Translation: A Systematic Analysis of Model Pipeline Optimisation Approaches and their Interactions arXiv:2607.10911v1 Announce Type: new Abstract: In the age of large language models, Natural Language to SQL (NL2SQL) translation remains an open problem with many useful applications. We explore interactions between several NL2SQL pipeline extensions to inspire development of… 18 TechCrunch — AI news-outlet 1mo ago Video generation startup PixVerse raises $439M, valuation soars past $2B Singapore-based video generation startup PixVerse closed a Series C extension on the strength of 15 million monthly active users, it said. 14 Vercel — AI dev-tools 1mo ago Vercel Plugin now available in VS Code and GitHub Copilot CLI The Vercel Plugin is now available in VS Code and the GitHub Copilot CLI. GitHub Copilot now has Vercel platform knowledge on demand, with skills for Next.js, AI SDK, Vercel Functions, and more. The Vercel plugin also helps Copilot stay up to date with the latest Vercel APIs and… 28 r/LocalLLaMA community 1mo ago Qwen3.6-35B-A3B-MTP built this website with Pi Coding Agent on my laptop I tested Unsloth's Qwen3.6-35B-A3B-MTP-GGUF:UD-Q4_K_XL with simple Pi Coding Agent. No skill or extensions used. My laptop have 32GB RAM and 8GB VRAM. Without context it runs around 40 t/s. When context getting bigger, speed can drop more than half. I made this website with one… 16 r/LocalLLaMA community 1mo ago I got Gemma 4 running directly inside Godot using only GDScript and Vulkan compute shaders I wanted to see if an LLM could run inside Godot without llama.cpp, Python, a server, or a GDExtension. It works. This Godot 4.7 project runs gemma-4-E2B-it-Q4_K_M.gguf locally. The model calculations run in Vulkan compute shaders, while GDScript handles GGUF loading,… 27 r/LocalLLaMA community 1mo ago Mellum2 with MTP? When JetBrains introduced Mellum 2, they advertised its latency being as low as Qwen2.5-Coder 7B. This was achieved via MTP. In the GGUFs they've published, I don't see layer resembling an MTP head, however. Is there some way to extract the MTP weights from safetensors directly?… 30 r/LocalLLaMA community 1mo ago LM Studio + Zoo + Qwen 3.6 issues I'm currently running an Unsloth quant of Qwen3.6-35B-A3B and I've managed to speed up my output to ~80 tk/s by offloading all experts to CPU. I'm on a Legion 7i laptop, 5080 with 16 GB VRAM. I mainly use the models for Zoo (formerly Roo) integration into VSCode. I'm running… 38 r/LocalLLaMA community 1mo ago Qwenthropic Hey guys, I've been running Qwen 3.6-27b locally on an RTX 3090 for a while now, and it's been genuinely great at solving software issues. However, life happened and I recently had to use Opus 4.8 alongside the Zed editor and the Claude Code agent. While I can definitely see a… 35 Hugging Face Daily Papers research 1mo ago Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE Abstract A novel zero-shot method called Jet-Long enables efficient long-context processing for large language models by dynamically adapting rescaling factors and utilizing a bifocal attention mechanism that maintains high performance across varying sequence lengths. Generated… 36 Page 1 of 4 · 151 articles Older →