News / #ide Tag Ide 178 articles archived under #ide · RSS Sign in to follow Hugging Face Daily Papers research 1mo ago Oxygen-TryOn: Fashion-Native Foundation Model for Any-item Virtual Try-On Abstract We present Oxygen-TryOn, a unified foundation model for any-item virtual try-on. Rather than repurposing a general-purpose image editor, Oxygen-TryOn is fashion-native, built for try-on through a dedicated data engine and try-on-specific training. Given one or more… 6 r/LocalLLaMA community 1mo ago NYT: Protect America’s lead in the A.I. race. “China is working hard to catch up, and the United States should take steps to keep its advantage. Most important, it should continue to prohibit American companies from selling the most advanced chips and equipment to China.” — The editorial board Ridiculous piece from the… 27 arXiv — NLP / Computation & Language research 1mo ago Demonstrating GenDB: Instance-Optimized and Customized Query Processing Code Generation via LLM Agents arXiv:2607.20630v1 Announce Type: cross Abstract: Traditional query processing engines require continuous development and extensions to support new techniques and user requirements, and in some cases, entirely new systems must be built from scratch. However, these engines are… 34 Hugging Face Daily Papers research 1mo ago Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning Abstract Reinforcement learning with verifiable rewards (RLVR) has substantially improved language-model reasoning, yet its extension to vision-language models remains constrained by the lack of training data that are simultaneously broad, exactly verifiable, and reproducible.… 36 arXiv — Machine Learning research 1mo ago Interpretable Fuzzy Rule-Based Regression Extension for Ex-Fuzzy Library arXiv:2607.20277v1 Announce Type: new Abstract: Machine learning models achieve high predictive accuracy in regression tasks, but their deployment in safety-critical and regulated domains requires interpretability. While fuzzy rule-based systems offer transparent, linguistically… 23 arXiv — NLP / Computation & Language research 1mo ago The Impact of Editorial Intervention on Detecting Native Language Traces arXiv:2605.10216v2 Announce Type: replace Abstract: Native Language Identification (NLI) is the task of determining an author's native language (L1) from their non-native writing. With the advent of human-AI co-authorship, learner texts are routinely corrected and rewritten by… 19 Vercel — AI dev-tools 1mo ago GitHub tools are now an installable eve extension You can now add GitHub tools to your eve agent as an extension . Add the package, drop one file in agent/extensions/ , and your agent gets all 42 tools with Vercel Connect auth, presets, and approval rules built in. Install @github-tools/eve-extension : Then register it from a… 25 arXiv — Machine Learning research 1mo ago A Self-Evolving Default Action for Cooperative Tasks with Continuous Action Space arXiv:2607.18597v1 Announce Type: new Abstract: Counterfactual credit assignment has proven effective in multi-agent reinforcement learning (MARL) for discrete action spaces, yet its extension to continuous-action cooperative tasks remains challenging. Existing methods that… 32 arXiv — NLP / Computation & Language research 1mo ago A Situational Speech Synthesizer for Yoruba: System Design, Phonological Rule Architecture, and Orthographic Extensions for Contour arXiv:2607.18317v1 Announce Type: cross Abstract: We present TTSYoruba, a rule-based concatenative diphone speech synthesizer for Yoruba, deployed at online as part of the YorubaName.com open dictionary of Yoruba personal names. The system takes tone-marked Yoruba text as input… 32 Vercel — AI dev-tools 1mo ago Extend eve agents with installable extensions You can now package tools, connections, skills, instructions, and hooks into extensions that any eve agent can import. Extensions can be published to package registries like npm, then installed, versioned, and upgraded like any other project dependency. A browser-use extension… 11 r/LocalLLaMA community 1mo ago pi 0.81.0 adds support for llama.cpp pi 0.81.0 now has integrated support for llama.cpp (llama-server router). https://pi.dev/docs/latest/llama-cpp This seems to be able to replace the huggingface/pi-llama extension and/or manually managing models in the config.   submitted by   /u/popoppypoppylovelove… 19 arXiv — Machine Learning research 1mo ago Periodic Bootstrap Thompson Sampling For Periodically Non-Stationary Bandit Problems arXiv:2607.16986v1 Announce Type: new Abstract: This paper introduces Periodic Bootstrap Thompson Sampling (PBTS), an innovative extension of the classic Thompson Sampling (TS) algorithm tailored for bandit problems with periodic non-stationarity. Conventional TS accumulates all… 38 r/LocalLLaMA community 1mo ago How did I do boys? Got it for 88k PHP (~$1,443). Owner was a video editor who recently upgraded to a Macbook Pro M5 48GB SPECIFICATIONS: AMD Ryzen 9 5950X (16 Cores / 32 Threads) ASUS ROG Strix X570-I Gaming (Mini-ITX) RTX 3090 24GB 64GB DDR4 RAM Kingston NV2 2TB NVMe SSD Thermaltake SFX-L 1000W… 25 arXiv — Machine Learning research 1mo ago qZACH-ViT: Quantization-Aware Intrinsic Explanations with Recursive Attribution-Stabilized Optimization arXiv:2607.15421v1 Announce Type: new Abstract: Compact medical-image classifiers need efficiency and interpretable evidence, yet these goals are often addressed separately. We introduce qZACH-ViT, a quantization-aware extension of the zero-token (CLS-token-free), position-free… 19 arXiv — Machine Learning research 1mo ago AV-JEPA: Extending LeJEPA to Audio-Visual Self-Supervised Learning arXiv:2607.15295v1 Announce Type: cross Abstract: We present AV-JEPA, an elegant multimodal extension of LeJEPA to audio-visual self-supervised learning. Using an early-fusion Vision Transformer and modality dropout as masking, the model is trained to align the embeddings of… 31 arXiv — Machine Learning research 1mo ago Counterfactuals for Feature-Weighted Clustering arXiv:2607.14719v1 Announce Type: new Abstract: Counterfactual explanations provide local, interpretable insight by identifying changes to an input that would alter its assigned outcome. Although well established in supervised learning, their extension to clustering is less… 12 Smol AI News news-outlet 1mo ago not much happened today **OpenAI's agent products** saw a **2.5x weekly usage growth** driven by **Codex + ChatGPT Work** and demand for **GPT-5.6 Sol**. JetBrains adopted Codex as a recommended agent, while LangChain enhanced tracing and observability across multiple tools. **PrismML released Bonsai… 36 arXiv — NLP / Computation & Language research 1mo ago A Stepwise Questioning Expert-Editor Multi-Agent Framework for Long-Document Summarization arXiv:2607.10390v1 Announce Type: new Abstract: Although large language models (LLMs) have shown promising potential in news summarization tasks, their performance on long-document summarization remains challenging as their length often exceeds the input limits. As the agent… 14 arXiv — NLP / Computation & Language research 1mo ago The Nuts and Bolts of Natural Language to SQL Translation: A Systematic Analysis of Model Pipeline Optimisation Approaches and their Interactions arXiv:2607.10911v1 Announce Type: new Abstract: In the age of large language models, Natural Language to SQL (NL2SQL) translation remains an open problem with many useful applications. We explore interactions between several NL2SQL pipeline extensions to inspire development of… 18 TechCrunch — AI news-outlet 1mo ago Video generation startup PixVerse raises $439M, valuation soars past $2B Singapore-based video generation startup PixVerse closed a Series C extension on the strength of 15 million monthly active users, it said. 14 Vercel — AI dev-tools 1mo ago Vercel Plugin now available in VS Code and GitHub Copilot CLI The Vercel Plugin is now available in VS Code and the GitHub Copilot CLI. GitHub Copilot now has Vercel platform knowledge on demand, with skills for Next.js, AI SDK, Vercel Functions, and more. The Vercel plugin also helps Copilot stay up to date with the latest Vercel APIs and… 28 r/LocalLLaMA community 1mo ago Qwen3.6-35B-A3B-MTP built this website with Pi Coding Agent on my laptop I tested Unsloth's Qwen3.6-35B-A3B-MTP-GGUF:UD-Q4_K_XL with simple Pi Coding Agent. No skill or extensions used. My laptop have 32GB RAM and 8GB VRAM. Without context it runs around 40 t/s. When context getting bigger, speed can drop more than half. I made this website with one… 16 r/LocalLLaMA community 1mo ago I got Gemma 4 running directly inside Godot using only GDScript and Vulkan compute shaders I wanted to see if an LLM could run inside Godot without llama.cpp, Python, a server, or a GDExtension. It works. This Godot 4.7 project runs gemma-4-E2B-it-Q4_K_M.gguf locally. The model calculations run in Vulkan compute shaders, while GDScript handles GGUF loading,… 27 r/LocalLLaMA community 1mo ago Mellum2 with MTP? When JetBrains introduced Mellum 2, they advertised its latency being as low as Qwen2.5-Coder 7B. This was achieved via MTP. In the GGUFs they've published, I don't see layer resembling an MTP head, however. Is there some way to extract the MTP weights from safetensors directly?… 30 r/LocalLLaMA community 1mo ago LM Studio + Zoo + Qwen 3.6 issues I'm currently running an Unsloth quant of Qwen3.6-35B-A3B and I've managed to speed up my output to ~80 tk/s by offloading all experts to CPU. I'm on a Legion 7i laptop, 5080 with 16 GB VRAM. I mainly use the models for Zoo (formerly Roo) integration into VSCode. I'm running… 38 r/LocalLLaMA community 1mo ago Qwenthropic Hey guys, I've been running Qwen 3.6-27b locally on an RTX 3090 for a while now, and it's been genuinely great at solving software issues. However, life happened and I recently had to use Opus 4.8 alongside the Zed editor and the Claude Code agent. While I can definitely see a… 35 Hugging Face Daily Papers research 1mo ago Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE Abstract A novel zero-shot method called Jet-Long enables efficient long-context processing for large language models by dynamically adapting rescaling factors and utilizing a bifocal attention mechanism that maintains high performance across varying sequence lengths. Generated… 36 arXiv — Machine Learning research 1mo ago Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE arXiv:2607.07740v1 Announce Type: new Abstract: Modern LLMs are increasingly deployed in long-context applications such as retrieval-augmented generation, repository-level coding, and agentic workflows whose accumulated reasoning and tool traces routinely push the input an order… 37 arXiv — NLP / Computation & Language research 1mo ago LEXIC: Lightweight Eye-tracking eXtension via Injected Complexity arXiv:2607.08152v1 Announce Type: new Abstract: On the recent EyeBench benchmark, predicting reading comprehension from eye movements exposes a stark gap: text-aware models using pretrained language models reach 56--63% AUROC, while gaze-only models operate at chance. We ask how… 8 TechCrunch — AI news-outlet 1mo ago OpenAI is shutting down Atlas, but its AI browser ambitions are still growing OpenAI is sunsetting its AI-powered browser after less than a year. But it's moving some agentic browsing features to its desktop app and a Chrome extension. 5 TechCrunch — AI news-outlet 1mo ago Paris-based AI voice startup Gradium raises $100M seed, backed by Nvidia The Paris-based ElevenLabs competitor, just announced a hefty seed extension round. 11 Hugging Face Daily Papers research 1mo ago Bibby AI: An Editor-Native Agentic Platform for Academic Research, Writing, and Publishing Abstract Bibby AI is an editor-native platform that consolidates the academic writing workflow into a unified Research-Write-Publish pipeline, streamlining literature management, citation insertion, and formatting through integrated tools and agents operating on document syntax… 14 arXiv — NLP / Computation & Language research 1mo ago The Role of Prompt Language and Translation-Theory-Driven Prompts in Large Language Models: A Case Study on Spanish-Chinese Journalistic Translation arXiv:2607.03160v1 Announce Type: new Abstract: This study examines how prompt language and translation theory-driven prompt design influence the quality of Spanish-Chinese journalistic translations generated by GPT-5.2. A parallel corpus of four editorials from El Pais was… 27 r/LocalLLaMA community 1mo ago Am I Expecting Too Much? I’ve been trying to get started with local coding help after finding Claude Code really useful but I’m just running into a ton of problems. On my RTX 5090, I’ve been running Qwen 3.6 27B UD_4 at 131K context, no KV quants, quant’d by Unsloth. I’ve been using Cline in VS Code, as… 5 r/LocalLLaMA community 1mo ago Using llama.cpp with pi A lot of people have their own setups using pi with llama.cpp. I have my own so thought I'll share. This is the extension I and deepseek wrote: https://github.com/am17an/pi-llama-server . It allows you to do two very simple things: auto detect a llama-server running and list the… 27 r/LocalLLaMA community 1mo ago GH Copilot’s BYOK Blocking for Inline Completion Makes No Sense. [THE FIX] GitHub Copilot allows Bring Your Own Key (BYOK) for its Chat window, but it explicitly blocks those exact same custom models from being used for inline code auto-completion. The official justification from the VS Code team is a supposed "lack of capable FIM (Fill-in-the-Middle)… 10 r/MachineLearning community 1mo ago I built a open source neural network shape validator [P] Built a visual editor that validates tensor shapes, counts params, estimates FLOPs/VRAM while you design. Catches incompatible residuals, mismatched Linear layers, all that before you waste GPU time. 63 ops. Proper shape inference. Exports PyTorch code that actually runs. URL-… 27 Hacker News — AI on Front Page community 1mo ago Wordgard: In-browser rich-text editor from the creator of ProseMirror Article URL: https://wordgard.net/ Comments URL: https://news.ycombinator.com/item?id=48772573 Points: 205 # Comments: 76 18 arXiv — Machine Learning research 1mo ago Gaming Consensus: Coordinated Manipulation in Crowdsourced Fact-Checking arXiv:2607.01824v1 Announce Type: new Abstract: Crowdsourced fact-checking systems have been adopted by major social media companies such as X, Meta, TikTok and Google with the aim of combating misleading information at scale without relying on centralized editorial control.… 15 r/LocalLLaMA community 1mo ago ZCode: New Agentic Code Editor from the Makers of GLM   submitted by   /u/johnnyApplePRNG [link]   [comments] 16 r/LocalLLaMA community 2mo ago Mellum2 local deployments Hey local community, I work at JetBrains with the team that trained Mellum2 models — 12B-2.5A LLMs. Those models are trained completely from scratch, targeting fast inference: our primary goal were H100/H200s prod deployments, but local deployments are good as well. We… 37 Hacker News — AI on Front Page community 2mo ago Show HN: OpenKnowledge – open source AI-first alternative to Obsidian/Notion Hi HN, Nick here. We’re launching OpenKnowledge ( https://openknowledge.ai/ ), a “what you see is what you get” markdown editor that has direct integrations with Claude, Codex, and other agents. Available as MacOS app or Web UI+CLI. Fully free/local and OSS. We built this… 20 Hacker News — AI on Front Page community 2mo ago LuaJIT 3.0 proposed syntax extensions Article URL: https://github.com/LuaJIT/LuaJIT/issues/1475 Comments URL: https://news.ycombinator.com/item?id=48667336 Points: 201 # Comments: 119 7 r/LocalLLaMA community 2mo ago SDXL running locally in the browser on WebGPU, open-source I needed simple local image generation without the usual setup. No virtual environments, no ComfyUI with a complex graph and installation as an exe. So i tried to push the whole thing into the browser and run it on WebGPU. It's a browser extension. You install it, then it loads… 13 Hacker News — AI on Front Page community 2mo ago Show HN: TikZ Editor – WYSIWYG editor for figures in LaTeX Hi all! TikZ is a widely-used LaTeX package for drawing figures in papers. It uses commands like \draw[->] (0,0) -- (1,2); to draw lines, shapes, text, etc. Academics usually code up their figures by hand, so there is lots of twiddling around with the coordinates and recompiling… 31 r/LocalLLaMA community 2mo ago Qwen code companion on vscode marketplace - thoughts I just came across this extension in vscode few days ago and tried to use with LM studio hosted models and it really is pretty good compared to `continue`, `kilo`, `cline`, `roo` like I felt without much tweaks, gets straight to the point, if any tweaks required u could do… 36 arXiv — Machine Learning research 2mo ago Latent Confounded Causal Discovery via Lie Bracket Geometry arXiv:2606.19610v1 Announce Type: new Abstract: Recent work on Kan-Do-Calculus (KDC) has established that the boundary between passive observation and active intervention in causal inference is a category-theoretic bi-adjunction, with interventions modeled by left Kan extensions… 16 arXiv — Machine Learning research 2mo ago Seeing Before Reasoning: Decoupling Perception and Reasoning for Shortcut-Resilient Multimodal On-Policy Self-Distillation arXiv:2606.19120v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a model on its own rollouts and uses a frozen copy to provide dense token-level targets conditioned on a reference target. This works well for LLM reasoning, but a direct extension to… 31 arXiv — Machine Learning research 2mo ago INDEQS: Informed Neural controlled Differential EQuationS arXiv:2606.19138v1 Announce Type: new Abstract: Neural Controlled Differential Equations (NCDE) provide a powerful continuous-time framework for forecasting time series, but standard graph-based extensions typically learn spatial structure purely from data, even in settings… 38 Hacker News — AI on Front Page community 2mo ago RFC 10008: The new HTTP Query Method Article URL: https://www.rfc-editor.org/info/rfc10008/ Comments URL: https://news.ycombinator.com/item?id=48568502 Points: 219 # Comments: 105 13 Page 2 of 4 · 178 articles ← Newer Older →