News / #security Tag Security 500 articles archived under #security · RSS Sign in to follow arXiv — Machine Learning research 20d ago Sharding Prevents LLM Oversight Failures and Adversarial Exploitation arXiv:2608.06422v1 Announce Type: new Abstract: Giving an LLM judge more compute does not necessarily make it check more requirements. When one call must return many verdicts, some decisions become weakly grounded in the evidence, even when that call receives the same token or… 13 arXiv — Machine Learning research 20d ago Diffusion LLMs as Targets and Adversaries: Mechanistic Safety Exploits arXiv:2608.07430v1 Announce Type: new Abstract: Diffusion Large Language Models (DLLMs) replace autoregressive next-token prediction with iterative parallel denoising, yet their internal safety mechanisms remain poorly understood. In this work, we investigate DLLMs both as… 36 arXiv — Machine Learning research 20d ago Cascade: Exploiting SLO-Aware latency budget for fair and high goodput LLM inference serving arXiv:2608.06557v1 Announce Type: cross Abstract: The reasoning and agentic capabilities of large language models have expanded the range of applications they support, from short interactive exchanges to long, compute-heavy requests. LLM serving platforms today define… 34 arXiv — NLP / Computation & Language research 20d ago Georeferencing Non-Gazetteered Place Names using Biological Specimen Records arXiv:2608.06884v1 Announce Type: new Abstract: Biological specimen records collected by natural history institutions constitute a rich source of temporal geographic knowledge, capturing biodiversity information about regional landscapes as they were recorded at different times.… 19 arXiv — NLP / Computation & Language research 20d ago Confirming Our Biases? Evaluating the Capabilities, Risks, and Societal Impact of Large Language Models arXiv:2608.06977v1 Announce Type: new Abstract: It is well established that large language models (LLMs) are sensitive to prompt framing, reflecting patterns in their training data or prior prompts. In this study, we investigate the extent to which LLMs reinforce users biases… 5 arXiv — NLP / Computation & Language research 20d ago Skaling: Chinchilla's Exponents Meet Kaplan's Coupling arXiv:2608.07222v1 Announce Type: new Abstract: Neural scaling laws are foundational for language model development, yet standard formulations systematically under- and overestimate loss at data-scarce and overtraining extremes. This failure originates in the underlying… 23 arXiv — NLP / Computation & Language research 20d ago LitTraceQA: A Benchmark for Multi-Stage Grounding and Verification in Scientific Question Answering arXiv:2608.07370v1 Announce Type: new Abstract: Scientific literature is increasingly used as a knowledge source for language models, retrieval-augmented generation systems, and research assistants, but answering research questions from papers requires more than fluent… 15 arXiv — NLP / Computation & Language research 20d ago CoinRAG: Contextualized Information Nugget KV Cache Reuse for Long-Context RAG arXiv:2608.07458v1 Announce Type: new Abstract: Recent optimization studies on Retrieval-Augmented Generation (RAG) have exploited chunk-level KV cache reuse to avoid processing long retrieved contexts for higher efficiency, while significant information redundancy and noise… 35 arXiv — NLP / Computation & Language research 20d ago StepJack: Benchmarking Computer-Use Agent Safety Against Multi-Step Indirect Prompt Injection arXiv:2608.06477v1 Announce Type: cross Abstract: Computer-use agents (CUAs) face a growing threat from indirect prompt injection, where adversarial instructions are planted in the environment such as web pages. In this paper, we introduce multi-step indirect prompt injection, a… 5 arXiv — NLP / Computation & Language research 20d ago Can MLLMs Decode the Creative Leap? Introducing C4 for Cross-Concept Understanding arXiv:2608.06501v1 Announce Type: cross Abstract: Creative capabilities of MLLMs matter in design, communication, education, and human--AI collaboration, yet remain difficult to evaluate because explicit targets and reward signals are scarce compared with accuracy-oriented… 34 arXiv — NLP / Computation & Language research 20d ago Let's Unlearn Stereotypes Before Decision-Making: Assessing the Impact of Intrinsic Bias Mitigation on Downstream Fairness in LLMs arXiv:2509.16462v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly used in high-stakes decision-making systems, where biased predictions can reinforce social and economic disparities. Although prior work has examined intrinsic representational bias… 8 arXiv — NLP / Computation & Language research 20d ago Kimi K2.5: Visual Agentic Intelligence arXiv:2602.02276v2 Announce Type: replace Abstract: We introduce Kimi K2.5, an open-source multimodal agentic model designed to advance general agentic intelligence. K2.5 emphasizes the joint optimization of text and vision so that two modalities enhance each other. This… 28 arXiv — NLP / Computation & Language research 20d ago Dependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low-Resource Languages arXiv:2605.02608v2 Announce Type: replace Abstract: Transformer-based models achieve state-of-the-art dependency parsing for high-resource languages, yet their advantage over simpler architectures in low-resource settings remains poorly understood. We evaluate four parsers---the… 32 Hugging Face official-blog 20d ago Meta is back with Muse Glimmer: local, agentic, multimodal, and open source Back to Articles a]:hidden"> Meta is back with Muse Glimmer: local, agentic, multimodal, and open source! Published August 10, 2026 Update on GitHub Upvote 4 Pedro Cuenca pcuenq merve merve ben burtenshaw burtenshaw Aritra Roy Gosthipaty ariG23498 Great news from the OGs of open… 34 Vercel — AI dev-tools 20d ago Simplified onboarding for deepsec deepsec , the open-source security review harness from Vercel, now lets you set up a repository and run its first security review with a single command. The init command now automates the standard setup process: creates the isolated .deepsec/ workspace, the only thing added to… 15 Simon Willison community 20d ago Quoting Claude Opus 5 system prompt Claude Fable 5 and Claude Mythos 5 were first released on June 9, 2026. On June 12, 2026, Anthropic suspended access to both models to comply with U.S. Department of Commerce export controls; the Department lifted those controls on June 30, 2026, and Anthropic restored access on… 11 Simon Willison community 20d ago Quoting Claude Opus 5 system prompt Claude Fable 5 and Claude Mythos 5 were first released on June 9, 2026. On June 12, 2026, Anthropic suspended access to both models to comply with U.S. Department of Commerce export controls; the Department lifted those controls on June 30, 2026, and Anthropic restored access on… 30 TechCrunch — AI news-outlet 20d ago Embattled hedge fund Situational Awareness invests $400M in chip startup Source Foundry The AI-focused hedge fund is still making some big bets. 10 r/LocalLLaMA community 20d ago Speculative decoding in a tools call paper : https://arxiv.org/html/2608.00814v1 source : https://x.com/i/status/2086505517640540587   submitted by   /u/Illustrious-Swim9663 [link]   [comments] 13 r/MachineLearning community 20d ago A Mechanistic Explanation of Prompt Injection (and why you should study roles) [R]   submitted by   /u/katxwoods [link]   [comments] 24 TechCrunch — AI news-outlet 21d ago Planned Amazon data center could become the biggest climate polluter in the U.S. As part of a planned Texas data center, Amazon is investing in an on-site power plant that could reportedly become the largest source of climate pollution in the United States. 9 Ars Technica — AI news-outlet 22d ago DeepMind’s hurricane breakthrough has surprised weather scientists Open source WeatherNext model can make accurate predictions with lower-resolution weather data. 6 r/MachineLearning community 22d ago What is currently considered the theoretically optimal quantization bit-width for LLMs? [D] I’m curious whether there is now a theoretical or empirical “sweet spot” for LLM quantization, preferably research done using open-source formats like GGUF Suppose you have a fixed memory/compute budget and can choose the model size freely. For example, instead of a smaller… 34 Don't Worry About the Vase community 22d ago OpenAI Trained Its Models For Months While Those Models Were Coordinating Exploits Via Message Boards How does the situation keep turning out to be worse than we know? 5 r/LocalLLaMA community 23d ago RTX 5090 Owner Built An Open-Source Tool That Shuts Down PC If It Detects The 12VHPWR Cable Drawing Too Much Power, But It Can Only Work On Specific GPUs GitHub : https://github.com/humza-khalid/12vhpwr-guard Reddit thread : https://www.reddit.com/r/nvidia/comments/1vglua1/i_built_a_free_open_source_tool_that_shuts_your/   submitted by   /u/pmttyji [link]   [comments] 9 Hugging Face Daily Papers research 23d ago MameLoshnLM: Yiddish Language Model and Evaluation Benchmark Abstract We present MameLoshnLM, the first open-source 8B-parameter language model built specifically for Yiddish. Despite Yiddish's rich textual tradition, its limited digital presence and the scarcity of reliable evaluation resources have constrained progress in Yiddish… 5 r/MachineLearning community 23d ago CIKM 2026 decisions [R] CIKM 2026 decisions will be announced today. The resource track outcomes have started going out. How did you go with CIKM 2026?   submitted by   /u/Happy-Hustler [link]   [comments] 36 Hugging Face Daily Papers research 23d ago Invisible Shortcuts: Why Vision Encoders Know Your Camera Abstract Deep vision models exploit shortcuts, relying on cues that correlate with supervision signals. Prior work has focused on visible biases, such as object-background or texture correlations. We identify a different source of shortcut learning: invisible metadata traces… 37 arXiv — Machine Learning research 23d ago Spectral Aliasing Pretext: A novel task for Self-Supervised fault diagnosis in rotating machinery arXiv:2608.05705v1 Announce Type: new Abstract: Deep learning is a new way for machinery fault diagnosis but requires extensive labeled data, a scarce resource in industrial settings. We propose Spectral Aliasing Pretext (SAP), a self-supervised learning method that pretrains… 19 arXiv — Machine Learning research 23d ago Learning to Rank Tensor Network Contraction Plans for GPU-Accelerated Quantum Circuit Simulation arXiv:2608.05819v1 Announce Type: new Abstract: Classical simulation remains essential for developing and validating quantum algorithms, but its cost grows rapidly with circuit size. Tensor-network contraction can reduce this cost by exploiting circuit structure, although its… 34 arXiv — Machine Learning research 23d ago Dynamic Graph Prompting via Topology-Routed Mixed-Curvature Experts arXiv:2608.06031v1 Announce Type: new Abstract: Dynamic graph prompting freezes a pre-trained temporal backbone and adapts it to label-scarce downstream tasks using lightweight prompts. However, existing methods operate within a single, fixed embedding space. In this work, we… 4 arXiv — Machine Learning research 23d ago SAGA: Score-Weighted Adaptive Generation Alignment for Low-Resource Nordic Language Models arXiv:2608.06179v1 Announce Type: new Abstract: Preference optimisation has proven effective for improving large language models but typically relies on costly human preference annotations. Extending these methods to morphologically rich, low-resource languages remains… 17 arXiv — Machine Learning research 23d ago RxnCLF: Contrastive Transformation-Aware Reaction Foundation Model for Improved Reactivity Prediction arXiv:2608.06259v1 Announce Type: new Abstract: Reaction yield prediction remains challenging because labeled data are scarce and reaction space is both combinatorially large and sparsely populated, limiting the generalization of existing reaction representations. String-,… 15 arXiv — NLP / Computation & Language research 23d ago Where Privacy Risk Lives in English-Source Multilingual RAG: A Stage-Decomposed Audit Across Five Query Languages arXiv:2608.05163v1 Announce Type: new Abstract: A common assumption holds that switching to a non-English language makes a multilingual RAG system easier to attack for personal information. We test this on an English-source synthetic-PII corpus with five query languages and a… 31 arXiv — NLP / Computation & Language research 23d ago A Study of ASR Adaptation and Representation Dimensionality Reduction in Persian Speech Emotion Recognition Using Whisper arXiv:2608.05165v1 Announce Type: new Abstract: Speech Emotion Recognition (SER) in low-resource languages remains a challenging problem due to limited labeled data. In this work, we study the use of Whisper for Persian SER with a particular focus on representation… 10 arXiv — NLP / Computation & Language research 23d ago Constraint-First Reasoning: A Training-Free Protocol for Exploiting Answer-Space Constraints in Mathematical Problem Solving arXiv:2608.05254v1 Announce Type: new Abstract: Large language models can derive a plausible mathematical object yet still violate explicit requirements--for example, by omitting a modular reduction, returning a non-integer, or using the wrong encoded answer form. We introduce… 8 arXiv — NLP / Computation & Language research 23d ago How to Recognize New Words: A Comparison Between Context Biasing Methods and Speech LLMs arXiv:2608.05759v1 Announce Type: new Abstract: Recognizing new and rare words - named entities, acronyms, domain specific special words, and other items scarce in training data - remains a key challenge for automatic speech recognition (ASR). We compare two strategies for this:… 19 arXiv — NLP / Computation & Language research 23d ago M$^3$R-Bench: A Unified Benchmark for Evidence-Grounded Multimodal Metaphor Understanding arXiv:2608.05817v1 Announce Type: new Abstract: Metaphor enables the understanding of abstract concepts through cross-domain mappings while conveying affective attitudes. In multimodal scenarios, visual and textual information jointly construct Target--Source mappings, requiring… 25 arXiv — NLP / Computation & Language research 23d ago MameLoshnLM: Yiddish Language Model and Evaluation Benchmark arXiv:2608.05850v1 Announce Type: new Abstract: We present MameLoshnLM, the first open-source 8B-parameter language model built specifically for Yiddish. Despite Yiddish's rich textual tradition, its limited digital presence and the scarcity of reliable evaluation resources have… 23 arXiv — NLP / Computation & Language research 23d ago RP-OPSD: Reasoning-Pivot-Guided On-Policy Self-Distillation for Multilingual Reasoning Transfer arXiv:2608.06347v1 Announce Type: new Abstract: Multilingual reasoning transfer is crucial for extending reasoning capabilities of large language models (LLMs) beyond high-resource languages. On-policy self-distillation (OPSD) and its variants have emerged as a promising… 32 arXiv — NLP / Computation & Language research 23d ago Learning Context-Free Grammars for Grammar-Constrained Decoding via Declarative Agentic Programming with Guarantees arXiv:2608.05493v1 Announce Type: cross Abstract: Language models (LMs) are increasingly used to interact with external services via programs written in domain-specific languages (DSLs). Unfortunately, since DSLs are often low-resource and esoteric, LMs frequently produce… 5 arXiv — NLP / Computation & Language research 23d ago EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents arXiv:2608.05519v1 Announce Type: cross Abstract: Agent benchmarks usually measure task completion and treat resource use as an auxiliary statistic. In deployment, however, the choice among a local lookup, broad search, composite research tool, stronger model, or human… 7 arXiv — NLP / Computation & Language research 23d ago The Vulnerability With No CVE: Managing Persistent Gaps Between Mandate and Authority in AI Coding Agents arXiv:2608.05884v1 Announce Type: cross Abstract: Existing guidance identifies excessive agency, excessive permission, weak task-bound authorization, and inadequate agent controls as important risks. Control frameworks also describe capabilities for constraining, authorizing,… 24 r/LocalLLaMA community 23d ago My issue with Artificial Analysis's 'intelligence index' I swear AA is not the bipartisan they so claim. An open source mode (Qwen 3.8 max) was number 1 on the agentic index, then they just so happen to launch "v4.1.1" of their index in which they just adjusted the weights of the gdpval and t3 banking so that it would be lower than… 13 r/LocalLLaMA community 23d ago Best open-source harnesses for combining cloud and local AI model orchestration? Looking for best current solutions for combining cloud models and local models seamlessly inside a harness' orchestration Edit: Right now, we don't have harnesses (that I'm aware of) that are blending local and cloud models to work together simultaneously to accomplish tasks set… 32 TechCrunch — AI news-outlet 24d ago Ex-Spotify employees raise $10M to bring the AI behind its recommendations to e-commerce The startup's platform predicts what product a shopper wants next, learn their general taste, and fine-tune continuously based on what they do in real time. 24 r/LocalLLaMA community 24d ago i just spent weeks rewriting my webUI from scratch, getting rid of all AI slop within the codebase and switching it over to a proper lightweight framework (alpine.js). i am now comfortable suggesting it as an alternative to openwebUI, librechat and the like! it is made for local… [Fully open source under GPL3, made from the ground up for use with local models, no subscriptions, no corporate backing] When i first started this, it was meant to be a fully lightweight, extremely modular alternative to openclaw, hermes and the like , and it still is! But i… 19 Hugging Face Daily Papers research 24d ago Agent Against Agent: An Agentic System for Automatic Prompt Injection Red Teaming Abstract Prompt injection poses significant security risks to LLM agents. Efficient and effective red-teaming is therefore critical, both for evaluating these risks and for collecting training data to improve defenses. Existing state-of-the-art prompt injection red-teaming… 25 arXiv — Machine Learning research 24d ago Spatiotemporal Graph Transformer for Traffic Intelligence in Edge Computing arXiv:2608.04075v1 Announce Type: new Abstract: Accurate traffic forecasting is essential for proactive resource management in edge computing, where service demand evolves dynamically across both space and time. In practical cellular edge systems, traffic exhibits strong spatial… 37 arXiv — Machine Learning research 24d ago Understanding Fault Tolerance of Adversarially Robust Pruned Models arXiv:2608.04173v1 Announce Type: new Abstract: Deep neural networks (DNNs) deployed on resource-constrained neuromorphic hardware face three concurrent challenges: the need for model compression through pruning, vulnerability to adversarial input perturbations, and… 29 Page 6 of 10 · 500 articles ← Newer Older →