News / #developer-tool Tag Developer Tool 500 articles archived under #developer-tool · RSS Sign in to follow arXiv — Machine Learning research 1mo ago The Planar Case of Thomas Positive Circuits Conjecture arXiv:2607.14143v1 Announce Type: cross Abstract: The notion of circuit refers to a cyclic oriented influence between the elements of a dynamical system. There are two classes of circuit: positive and negative. R. Thomas conjectured that a necessary condition of multi… 21 arXiv — NLP / Computation & Language research 1mo ago ReportMedSAM: Guiding Segmentation Through Radiology Reports arXiv:2607.14116v1 Announce Type: new Abstract: Free-form radiology reports contain rich clinical descriptions, yet converting them for reliable segmentation remains challenging due to the inherent variability of natural language. Existing pipelines often rely on predefined… 32 arXiv — NLP / Computation & Language research 1mo ago MamaBench: Benchmarking LLM Robustness in Maternal and Child Health Diagnosis through Counterfactual Clinical Perturbation arXiv:2607.14385v1 Announce Type: new Abstract: Large language models achieve strong scores on medical benchmarks, yet these benchmarks evaluate each question in isolation, providing no measure of whether a system can distinguish clinically similar presentations requiring… 11 arXiv — NLP / Computation & Language research 1mo ago MonteRET: AI Agent Enhancing Multimodal LLMs with Multi-granularity Knowledge Retrieval for Chest CT Report Generation arXiv:2607.14264v1 Announce Type: cross Abstract: Automated chest CT report generation remains challenging because clinically faithful reporting requires both whole-volume understanding and accurate description of localized anatomical findings. Here we developed and… 11 arXiv — NLP / Computation & Language research 1mo ago MedFailBench: A Clinician-Built Open-Source Benchmark for Medical AI Safety Boundary Inspection arXiv:2607.15166v1 Announce Type: cross Abstract: Most medical AI benchmarks measure whether a model knows the correct answer. MedFailBench asks a different question: which safety boundary failed? We present a clinician-built synthetic benchmark and failure atlas that labels… 27 Vercel — AI dev-tools 1mo ago How Searchable ships customer-requested features in 30 minutes on Vercel Searchable on Vercel 5x increase in development velocity 100+ billion tokens processed Customer-requested features shipped in as little as 30 minutes Zero model SDK implementation or API key rotation with AI Gateway Searchable helps brands track and improve how they appear… 29 Vercel — AI dev-tools 1mo ago GLM 5.2 is 35% off via Novita on AI Gateway GLM 5.2 is 35% off on AI Gateway through July 24 when routed through Novita. To get the discounted rate, set the model to zai/glm-5.2 in the AI SDK and route requests through Novita: After July 24, the model stays available at standard provider rates with no markup. Try GLM 5.2… 24 Vercel — AI dev-tools 1mo ago Chat SDK adds native Slack agent support You can now build native Slack agents with Chat SDK's Slack adapter . The adapter supports the full Slack agent messaging experience, from agent conversations in the Messages tab to suggested prompts, rotating status messages, token-by-token streamed replies, and native feedback… 33 Vercel — AI dev-tools 1mo ago Vercel Plugin now available in Kimi Code CLI The Vercel Plugin is now available in the Kimi Code CLI . Kimi Code can now draw on Vercel platform knowledge on demand, with skills for Next.js, AI SDK, Vercel Functions, and more. The Vercel Plugin also helps Kimi Code stay up to date with the latest Vercel APIs and… 28 TechCrunch — AI news-outlet 1mo ago Yes, you can now order DoorDash from the command line DoorDash is opening a limited beta of dd-cli, a command-line tool that lets developers and AI agents search stores, build carts, and place orders from the terminal, marking another step toward software designed for AI agents instead of just humans. 24 arXiv — Machine Learning research 1mo ago OrDA: Orthogonal Disentanglement of Access Habits Framework for Homepage Marketing Block Recommendations arXiv:2607.13420v1 Announce Type: new Abstract: Clicks on homepage marketing blocks are driven by a dual-mechanism of content interest and access habits. However, habitual clicks often create Pseudo-Positives in marketing slots, where position advantage masks mediocre content… 37 arXiv — Machine Learning research 1mo ago The Hyperspherical Geometry of CLIP Latent Space: A Semantic Mixture Model arXiv:2607.13660v1 Announce Type: new Abstract: Contrastive Language-Image Pretraining (CLIP) representations form a semantic embedding space governed by cosine similarity, reflecting an intrinsic hyperspherical geometry. However, existing probabilistic interpretations typically… 28 arXiv — Machine Learning research 1mo ago Improving Wind and Solar Power Prediction with Efficient Wrapper-based Feature Selection: An Empirical Study arXiv:2607.14024v1 Announce Type: new Abstract: With rising global energy demand and growing awareness of climate change and its impacts, the share of renewable energies in the global energy mix continues to grow. Unlike conventional power generation, the output of renewable… 19 Simon Willison community 1mo ago Mermaid to Unicode box art (grok-mermaid) Tool: Mermaid to Unicode box art (grok-mermaid) While exploring the codebase for the newly open-sourced Grok CLI coding agent I came across xai-grok-markdown/src/mermaid.rs , a "self-contained terminal renderer for Mermaid diagrams" written in Rust. I figured it would be fun to… 14 Simon Willison community 1mo ago xai-org/grok-build, now open source xai-org/grok-build, now open source xAI's grok CLI tool faced severe community backlash yesterday when it became apparent that running the command in a directory could upload that entire directory to xAI's Google Cloud buckets. One user reported running it in their home… 26 llama.cpp releases dev-tools 1mo ago b10031 tokenize : drop --stdin mutual-exclusion check ( #25672 ) match cli and completion, which don't enforce it macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu arm64 (CPU)… 13 r/LocalLLaMA community 1mo ago How to run Prism Bonsai 27B According to their GitHub , the file needed is the -Q2_0.gguf version, which requires compiling their version of llama.cpp . Easiest way to download is to use huggingface-cli if you have it available: hf download prism-ml/Ternary-Bonsai-27B-gguf --include "*-Q2_0.gguf*" If not,… 6 arXiv — Machine Learning research 1mo ago Graph-Constrained Policy Learning for Extreme Clinical Code Prediction arXiv:2607.11954v1 Announce Type: new Abstract: Clinical code prediction maps unstructured discharge summaries to ICD-10-CM leaf codes in a large, sparse, and deeply hierarchical label space. Most systems treat the task as flat multi-label classification, scoring codes… 35 arXiv — Machine Learning research 1mo ago GenDiff: A Dose and Anatomy Aware Diffusion Model with Structural Prior Refinement for Low-Dose CT Reconstruction and Generalization arXiv:2607.11941v1 Announce Type: cross Abstract: Computed tomography (CT) is a critical imaging modality for clinical diagnosis, but reducing radiation dose inevitably introduces severe noise and structured artifacts that degrade image quality. Existing deep learning-based… 31 arXiv — NLP / Computation & Language research 1mo ago Agentic systems for breast cancer treatment recommendations arXiv:2607.12051v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being explored for clinical decision support, but their reliability in complex oncology treatment planning remains unclear. We evaluated agentic LLM systems for breast cancer treatment… 28 arXiv — NLP / Computation & Language research 1mo ago An Empirical Analysis of Continual Learning for Heterogeneous Medical Visual Question Answering arXiv:2607.12048v1 Announce Type: cross Abstract: Deploying medical visual question answering (MedVQA) systems in real-world clinical settings requires models that adapt to new clinical tasks without forgetting previously acquired knowledge. Continual learning (CL) provides a… 36 Vercel — AI dev-tools 1mo ago How Speechify serves 500,000 dynamic pages to 60 million users on Vercel Speechify on Vercel 500,000+ pages served across 40+ languages Cut costs 50% by auto-scaling with Fluid compute Zero user impact on bad deploys with Instant Rollbacks Speechify started as a tool for people with dyslexia. Cliff Weitzman, Founder & CEO, built it because reading… 31 Vercel — AI dev-tools 1mo ago Chat SDK adds Discord Components V2 support You can now send Discord bot messages with Components V2 , an opt-in layout system that treats text, images, files, and buttons as flexible components you can arrange in any order. Set the adapter's contentFormat option to ComponentsV2 and cards render with native containers,… 4 Vercel — AI dev-tools 1mo ago Vercel Workflows trace viewer now has a minimap The trace viewer for Vercel Workflows now has a minimap that shows the entire run in one view. Drag the viewport to pan, resize it to zoom, or drag across the minimap to select a new section. Click anywhere on the minimap to jump to that point, or scroll and use arrow keys to… 7 Vercel — AI dev-tools 1mo ago Faster, predictable project linking in the Vercel CLI The Vercel CLI now resolves your team before discovering projects, then searches for projects only in that team instead of sweeping every team you belong to. Linking is faster and more predictable, and every command that establishes a link ( vercel link , deploy , pull , dev ,… 27 Vercel — AI dev-tools 1mo ago Vercel Connect support in GitHub Tools GitHub Tools now has first-class support for Vercel Connect through the new @github-tools/sdk/connect subpath. Instead of storing a long-lived personal access token, your agent mints short-lived, scoped GitHub tokens at runtime from a connector. There is no secret to store,… 38 Simon Willison community 1mo ago simonw/pedalican simonw/pedalican Clearly I wasn't paying attention when these were first announced back in May, but today I accidentally activated a "pet" in Codex Desktop - a little animated robot, reminiscent of Clippy - and then learned you can create your own. So I did, and now I have a… 35 arXiv — Machine Learning research 1mo ago FedCausal-Dyn: A Causal-Dynamic Paradigm for Federated Learning under Dynamic Feature Drift arXiv:2607.09695v1 Announce Type: new Abstract: This paper addresses the challenging problem of dynamic feature drift in federated learning, where data distributions evolve across clients and over time -- a common scenario in real-world applications like financial technology.… 18 arXiv — Machine Learning research 1mo ago Mitigating Early Training Collapse in CTR Models arXiv:2607.09696v1 Announce Type: new Abstract: Deep neural models for click-through rate prediction often exhibit a sharp decline in validation performance immediately after the first training epoch despite continued improvement in training loss. This instability restricts… 29 arXiv — Machine Learning research 1mo ago Manifold Constrained Tabular Deep Neural Networks arXiv:2607.09710v1 Announce Type: new Abstract: Tabular classification is often governed by local, condition-triggered rules rather than smooth global patterns. However, tabular deep neural networks (DNNs) are typically built upon Euclidean representations that favor smooth… 16 arXiv — Machine Learning research 1mo ago Multimodal Routing for Interpretable, Robust, and Auditable Clinical Prediction arXiv:2607.09982v1 Announce Type: new Abstract: Electronic health record (EHR) data are inherently multimodal, and leveraging multiple modalities can improve predictive performance. However, most existing approaches rely on deep fusion, which obscures how individual modalities… 25 arXiv — Machine Learning research 1mo ago Vilya-1: An all-atom foundation model for macrocycle structure prediction and design arXiv:2607.09998v1 Announce Type: new Abstract: Macrocyclic peptides are an increasingly important therapeutic modality, but existing computational methods for modeling their structures and properties are limited in scope and do not generalize well across the synthetically… 32 arXiv — Machine Learning research 1mo ago Beyond Euclidean Clipping: Overcoming Exploration Collapse in LLM RL via Riemannian Isometric Policy Optimization arXiv:2607.10169v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a dominant paradigm for enhancing LLMs' reasoning capabilities. However, RL algorithms with PPO-Clip are inherently limited by exploration collapse. Subsequent works remain primarily heuristic… 21 arXiv — Machine Learning research 1mo ago Pitfalls of Administrative Censoring in Survival Models with Time-Indexed Inputs arXiv:2607.10466v1 Announce Type: new Abstract: Survival models can model time-to-event outcomes using partially observed data. They are widely used in clinical prediction, including cancer risk, disease progression, treatment response, and mortality. Recent models often rely on… 38 arXiv — Machine Learning research 1mo ago On the modality gap and the contrastive loss in multi-modal representation learning arXiv:2607.10698v1 Announce Type: new Abstract: We study the modality gap in CLIP-style dual-encoder contrastive learning, where image and text embeddings remain misaligned despite being trained in a shared space. We argue that the gap is induced by a failure of the InfoNCE… 37 arXiv — Machine Learning research 1mo ago Policy-Driven CT-Agent: Modeling Phase-Aware Diagnostic Control for Clinically Consistent CT Reasoning arXiv:2607.10748v1 Announce Type: new Abstract: Computed Tomography (CT) diagnosis often relies on dynamic selection of imaging phases, such as non-contrast, arterial, or venous phases, based on preliminary findings, clinical suspicion, and diagnostic guidelines. This phase-wise… 27 arXiv — NLP / Computation & Language research 1mo ago CLIR-Bench: Benchmarking Multimodal Question Answering over Irregular Clinical Time Series arXiv:2607.09880v1 Announce Type: new Abstract: Clinical time series are central to patient monitoring, risk assessment, and clinical decision support. However, they are often sparse, irregularly sampled, and asynchronous, making it difficult for models to identify the temporal… 5 arXiv — NLP / Computation & Language research 1mo ago Faithful by Design: Evaluating and Improving LLM-Generated Clinical Trial Summaries for Multi-Stakeholder Audiences arXiv:2607.09932v1 Announce Type: new Abstract: Large language models are increasingly used to summarize clinical trial results for healthcare providers, patients, and payers, but their tendency to hallucinate poses significant risks in this high-stakes context. This study… 37 Vercel — AI dev-tools 1mo ago Chat SDK adds X adapter support Chat SDK now supports X (Twitter), extending its single-codebase approach to Slack, Discord, GitHub, Teams, Telegram, and WhatsApp with the new X adapter . Teams can build bots that reply to public mentions and hold direct message conversations through the X API v2 and the X… 15 Vercel — AI dev-tools 1mo ago Vercel Plugin now available in VS Code and GitHub Copilot CLI The Vercel Plugin is now available in VS Code and the GitHub Copilot CLI. GitHub Copilot now has Vercel platform knowledge on demand, with skills for Next.js, AI SDK, Vercel Functions, and more. The Vercel plugin also helps Copilot stay up to date with the latest Vercel APIs and… 28 Hacker News — AI on Front Page community 1mo ago Climate.gov was destroyed. Open data saved it Article URL: https://werd.io/climate-gov-was-destroyed-open-data-saved-it/ Comments URL: https://news.ycombinator.com/item?id=48897945 Points: 337 # Comments: 135 33 Hacker News — AI on Front Page community 1mo ago Former NOAA employees built Climate.us to preserve climate data and resources Article URL: https://19thnews.org/2026/07/noaa-climate-data-website/ Comments URL: https://news.ycombinator.com/item?id=48897945 Points: 408 # Comments: 161 35 MIT News — AI research 1mo ago How MIT students are helping to prevent cyberattacks Students from the MIT Cybersecurity Clinic help local governments and other vulnerable organizations defend against digital threats. 28 Hugging Face Daily Papers research 1mo ago MedPMC: A Systematic Framework for Scaling High-Fidelity Medical Multimodal Data for Foundation Models Abstract Medicine is inherently multimodal, requiring clinicians to synthesize information across diverse data streams. Yet the development of multimodal foundation models is constrained by limited access to large-scale, high-quality clinical data. Although PubMed Central (PMC)… 21 arXiv — Machine Learning research 1mo ago HERO: A Heterogeneity-Aware Benchmark Library for Federated Continual Learning arXiv:2607.08784v1 Announce Type: new Abstract: Federated continual learning (FCL) evaluates how distributed clients learn from changing data streams while retaining previously learned knowledge. Existing evaluations are difficult to compare because they often change datasets,… 35 arXiv — Machine Learning research 1mo ago EHR-MPC: Inference-Time Control for Sepsis Treatment with Generative Patient Digital Twins arXiv:2607.08793v1 Announce Type: cross Abstract: Sepsis is a leading cause of mortality, yet optimal treatment policies remain contested. Existing reinforcement learning (RL) approaches learn fixed strategies for sepsis treatment, limiting adaptability to changing clinical… 7 arXiv — Machine Learning research 1mo ago From Classification to Localization and Clinical Validation: Large-Scale Development of a Deep Learning System for Thoracic Disease Detection on Chest Radiographs in Thailand arXiv:2607.09305v1 Announce Type: cross Abstract: Chest radiography (CXR) remains the most widely used thoracic imaging modality, yet expert interpretation is constrained by a severe shortage of radiologists in Thailand and across Southeast Asia. Local adaptation of deep… 35 arXiv — NLP / Computation & Language research 1mo ago Deceptive Grounding: Entity Attribution Failure in Clinical Retrieval-Augmented Generation arXiv:2607.09349v1 Announce Type: new Abstract: Retrieval-augmented generation evaluation checks whether model claims are factually grounded in retrieved documents. It does not check whether retrieved evidence is attributed to the correct entity. A clinical RAG response can pass… 38 arXiv — NLP / Computation & Language research 1mo ago MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation arXiv:2607.09142v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in online medical consultation, yet existing benchmarks remain poorly aligned with real clinical practice. Many rely on synthetic conversations or patient simulators, omit… 30 arXiv — NLP / Computation & Language research 1mo ago Evaluating Retrieval-Augmented Generation vs. Long-Context Input for Clinical Reasoning over EHRs arXiv:2508.14817v2 Announce Type: replace Abstract: Objective: To evaluate whether retrieval-augmented generation (RAG) can serve as an efficient alternative to long-context prompting for clinical reasoning over electronic health records (EHRs). Methods: We defined three… 27 Page 9 of 10 · 500 articles ← Newer Older →