News / #security Tag Security 500 articles archived under #security · RSS Sign in to follow Hugging Face Daily Papers research 1mo ago Mapping CVEs to MITRE ATT&CK Techniques: A Curated Gold-Set Classifier and the Limits of LLM-Assisted Label Expansion Abstract We present a reproducible pipeline for mapping Common Vulnerabilities and Exposures (CVEs) to MITRE ATT&CK Enterprise techniques from free-text vulnerability descriptions. Rather than relying on the CWE->CAPEC->ATT&CK derivation chain, whose table-expansion artifacts we… 14 arXiv — NLP / Computation & Language research 1mo ago Where Steering Signals Come From: Activation Source Selection in Activation Steering arXiv:2607.25270v1 Announce Type: new Abstract: Activation steering controls language models by adding vectors or features to hidden states at inference time, but the upstream source of these steering signals is often treated as a secondary detail. We study this source choice as… 25 arXiv — NLP / Computation & Language research 1mo ago IRIS: Reusable Identity Representations from Frozen LLMs for Entity Alignment arXiv:2607.25579v1 Announce Type: new Abstract: Entity alignment (EA) identifies entities across knowledge graphs (KGs) that refer to the same real-world object. Conventional EA methods mainly exploit explicit graph structures and textual fields, which often provide insufficient… 7 arXiv — NLP / Computation & Language research 1mo ago The Effect of Text Chunk Size on Retrieval-Augmented Generation Performance arXiv:2607.24767v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems have emerged as a powerful process for allowing large language models (LLMs) to retrieve relevant information to use as source material during text generation. A critical yet… 11 arXiv — NLP / Computation & Language research 1mo ago On the Use of LLMs for Specialised Terminology: A Good Alternative to Corpora? arXiv:2607.24784v1 Announce Type: cross Abstract: Specialised translation relies on the use of documentary and terminological resources, including corpora. These resources are particularly useful for terminology. However, their compilation and exploitation have several… 6 arXiv — NLP / Computation & Language research 1mo ago Ranked by Position: Order Sensitivity as an Exploitable Attack Surface in LLM Listwise Recommenders arXiv:2607.24869v1 Announce Type: cross Abstract: Large language models (LLMs) used as listwise rerankers in recommendation systems suffer from position bias when serializing candidate sets into prompts. We show this order sensitivity creates an exploitable attack surface: an… 15 arXiv — NLP / Computation & Language research 1mo ago Building Large-Scale English-Romanian Literary Translation Resources with Open Models arXiv:2509.07829v4 Announce Type: replace Abstract: Literary translation has recently gained attention as a distinct and complex task in machine translation research, yet translation by small open models remains an open problem, particularly for low-resource languages such as… 12 r/LocalLLaMA community 1mo ago Nvidia is expected to raise GeForce RTX GPU prices again by up to 30%   submitted by   /u/ab2377 [link]   [comments] 21 r/LocalLLaMA community 1mo ago Agenta: an open-source Claude Cowork alternative where you can use self-hosted models (and any harness) Hey r/LocalLLaMA, I’m Mahmoud from Agenta. We built a self-hosted, more flexible, alternative to Claude Cowork . This short video shows how it works. I use it to build AI coworkers for my startup, like the marketing agent in the video. Or AI automations, like a daily report to… 33 Ars Technica — AI news-outlet 1mo ago We now have a better understanding how OpenAI hacked into Hugging Face 10 days passed from OpenAI models exploiting JFrog Artifactory 0-day to release of a patch. 33 r/MachineLearning community 1mo ago NeurIPS-side prompt injection triggering ethics reviewers? [D] Does anyone experience a similar story that some reviewers reporting ethical issue due to NeurIPS-side prompt injection for catching LLM-reviewers? Even ethics reviewers were not informed about this conference-side manipulation…   submitted by   /u/dontknowwhattoplay… 19 TechCrunch — AI news-outlet 1mo ago Fish Audio raises $50M seed to build AI voice models for creators and enterprises Since launching last year, the startup today has more than 8 million people using the open-source or hosted version of its models, and now generates annual recurring revenue of $21 million. 4 r/MachineLearning community 1mo ago NeurIPS 2026 AI-generated reviews [D] I'm really confused about what the point of the prompt injection was (speaking as an author). Is it just a study? I would really prefer that they took action against the AI-generated reviews. Obviously, we cannot assume that the reviewers were copy-pasting the output from the… 5 arXiv — NLP / Computation & Language research 1mo ago Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B arXiv:2607.22545v1 Announce Type: cross Abstract: Deploying large language models in financial-services and agentic settings requires safety classifiers that simultaneously handle prompt injection, regulatory compliance, and general harm, a combination no existing open guardrail… 21 arXiv — Machine Learning research 1mo ago Directional Influence Function: Estimating Training Data Influence in Constrained Learning arXiv:2607.23388v1 Announce Type: new Abstract: As constrained learning becomes increasingly common, models are trained under explicit feasibility requirements to enforce fairness, safety, robustness, regulariza- tion, and physics or logic constraints. Understanding how training… 25 arXiv — Machine Learning research 1mo ago When Can Depth Replace Precision? A Resource Theory of Quantized Neural Computation arXiv:2607.23390v1 Announce Type: new Abstract: When can additional low-bit residual computation replace missing numerical precision for a fixed input-output map? We model a quantized residual system over a fixed horizon as a pure schedule selecting fields from a declared… 4 arXiv — Machine Learning research 1mo ago Generalization bounds and sample complexity for remaining useful life prediction from complete degradation trajectories arXiv:2607.23454v1 Announce Type: new Abstract: Data-driven remaining useful life (RUL) prediction requires complete degradation trajectories for training, yet such run-to-failure data are scarce and expensive. Practitioners currently lack principled guidance on how many failure… 24 arXiv — Machine Learning research 1mo ago Extreme Volatility Warning under Label Scarcity via Multi-Source Anomaly Fusion arXiv:2607.23682v1 Announce Type: new Abstract: Early warning of extreme market volatility is central to financial risk management, but actionable events are rare, nonstationary, and often triggered by exogenous information shocks. In our CSI~300 setting, only $\sim$80 positive… 38 arXiv — NLP / Computation & Language research 1mo ago Explaining GAND: A Resource on Gender-Ambiguous Natural Data & Contrastive Attribution arXiv:2607.22546v1 Announce Type: new Abstract: Machine translation (MT) systems continue to produce gender-biased translations. In a time where self-expression is paramount, mistranslations based on default behaviour and stereotyping can lead to harm for users of these systems.… 11 arXiv — NLP / Computation & Language research 1mo ago PatiGonit22K: A Comprehensive Dataset for Solving Complex Bengali MWPs arXiv:2607.22859v1 Announce Type: new Abstract: Mathematical Word Problems (MWPs) are an important benchmark for evaluating natural language understanding and quantitative reasoning. Despite recent progress in high resource languages, Bengali remains underexplored due to the… 24 arXiv — NLP / Computation & Language research 1mo ago Simple Language Normalization Wins: Cross-Lingual Speaker Verification for the TidyVoice 2026 Challenge arXiv:2607.22923v1 Announce Type: new Abstract: Cross-lingual mismatch remains a key source of overall degradation in modern speaker verification. The TidyVoice2026 Challenge targets this setting with text-independent verification, comprising 3,666 training and 808 development… 29 arXiv — NLP / Computation & Language research 1mo ago BERT-based Models vs. Large Language Models for Low-Resource Named Entity Recognition: A Comparative Study on Marathi arXiv:2607.23344v1 Announce Type: new Abstract: Named Entity Recognition (NER) for low-resource languages such as Marathi remains a challenging task due to limited annotated resources and linguistic complexity. Although recent Large Language Models (LLMs) have demonstrated… 31 arXiv — NLP / Computation & Language research 1mo ago Guiding Language Models to Be More Empathetic: Culturally Sensitive Mental Health Advice Generation Through Human-LLM Collaboration arXiv:2607.23538v1 Announce Type: new Abstract: Despite recent advances in large language models (LLMs), their ability to generate empathetic mental health counseling responses in low-resource languages remains largely unexplored. To address this gap, we curate 625 authentic… 9 arXiv — NLP / Computation & Language research 1mo ago Pointer-Augmented Autoregressive Generation of Patent Claims with Joint Topology and Content Decoding arXiv:2607.24040v1 Announce Type: new Abstract: Autoregressive decoders emit flat token sequences and cannot enforce hierarchical constraints across output segments, a limitation that becomes acute in patent claim generation, where a claim set forms a dependency forest whose… 38 arXiv — NLP / Computation & Language research 1mo ago BioSentinel at EXIST 2026: Soft-Label Optimization with XLM-RoBERTa for Sexism Intent Classification in Memes arXiv:2607.24137v1 Announce Type: new Abstract: This paper describes the BioSentinel team's participation in EXIST 2026 Task 2.2: Source Intention in Memes, part of the CLEF 2026 evaluation campaign. The task requires classifying the communicative intent behind memes as direct,… 8 arXiv — NLP / Computation & Language research 1mo ago CAGE: Cognitive Attribution Graphs for Faithful Inline Citation Generation in Long-Form Question Answering arXiv:2607.24236v1 Announce Type: new Abstract: Long-form question answering increasingly relies on retrieved evidence to make LLM outputs verifiable, with inline citations tracing claims to source documents. However, existing systems often attach citations that are topically… 12 arXiv — NLP / Computation & Language research 1mo ago Systematic Analysis of Large Language Models and Transformer-Based Machine Translation for English-Tamil and Tamil-English Across Diverse Datasets arXiv:2607.24515v1 Announce Type: new Abstract: The challenge of Machine Translation for low resource languages such as Tamil is primarily caused by the restricted amount of parallel data for these languages, as well as their substantial amount of domain variation and… 14 arXiv — NLP / Computation & Language research 1mo ago Source-Aware Reranking for Retrieval-Augmented Generation: A Reliability Prior Approach arXiv:2607.22584v1 Announce Type: cross Abstract: Standard Retrieval-Augmented Generation pipelines rank retrieved documents by semantic similarity alone, without accounting for source provenance or credibility. This work evaluates a simple and interpretable modification to RAG… 5 Vercel — AI dev-tools 1mo ago Vercel Sandbox supports forking Vercel Sandbox now supports forking with Sandbox.fork() . The fork starts from the source's current snapshot and inherits its config and environment variables. If the source is running, it forks the latest saved state, not the live in-memory state. If the source has no snapshot,… 14 Hugging Face Daily Papers research 1mo ago Interactive Training 2: Auditable Control Plane for Live Model Training Abstract Experiment trackers show how training is progressing, but changing a live run still usually requires trainer-specific code. We present Interactive Training 2, an open-source control plane for steering training through a shared protocol. Training applications declare… 26 TechCrunch — AI news-outlet 1mo ago OpenAI’s Hugging Face breach has reignited the debate over alignment and control OpenAI's Hugging Face breach has reignited debate over AI alignment and control, exposing competing views on whether increasingly capable AI should be better aligned, better contained, or both. 14 NVIDIA Developer Blog official-blog 1mo ago NVIDIA Ising Enables Fully Automated Quantum Computer Calibration with Enhanced In-Context Learning NVIDIA Ising Calibration is an open source vision language model (VLM) designed to interpret diagnostic outputs from quantum processors and determine how they... 9 r/LocalLLaMA community 1mo ago Nvidia CEO Jensen Huang defends Open Source AI by saying distillation is fundamental to learning Nvidia CEO Jensen Huang “Distillation - learning from AI, learning from other people, and learning from other sources of knowledge, is fundamental to intelligence. We are constantly learning from one another. AI also has to learn from something.” Since using AI Desktop 98 , I… 23 r/LocalLLaMA community 1mo ago The entire tech industry (save for Anthropic) has come out in favor of open source AI. So what happens next? Will Anthropic change its lobbying efforts? Not likely. Now the gaslighting begins: “Nobody is trying to ban open source.”   submitted by   /u/SignificantLegs [link]   [comments] 34 r/LocalLLaMA community 1mo ago Meta has confirmed that it will release an open source model in the future https://preview.redd.it/k97l56d8ypfh1.png?width=606&format=png&auto=webp&s=2a2e2156bea56b25f5709e8f1df2bf82525eb089 https://x.com/alexandr_wang/status/2081501627836661928?s=20   submitted by   /u/External_Mood4719 [link]   [comments] 38 Smol AI News news-outlet 1mo ago not much happened today **Moonshot** released the **Kimi K3** open-weights model, a **2.8T-parameter MoE** with **104B active parameters**, **896 experts**, and **1M-token context** featuring native visual understanding. The release includes open-source infrastructure like **FlashKDA**, **MoonEP**, and… 33 arXiv — Machine Learning research 1mo ago Physiological Signals as a Forensic Modality for Talking-Face Deepfake Detection arXiv:2607.21776v1 Announce Type: new Abstract: Talking-face (TF) deepfake generation synthesizes photore- alistic facial video from a static source image and an au- dio signal, producing forgeries that current image-based detectors consistently fail to identify. Unlike… 9 arXiv — NLP / Computation & Language research 1mo ago LeAct: Learning to Reason from Expert Actions arXiv:2607.21856v1 Announce Type: cross Abstract: Modern reasoning models depend on reasoning data, today sourced from human annotations or distilled from stronger LLMs. However, a rich and largely untapped source of supervision lies in expert systems (e.g., game engines,… 17 arXiv — Machine Learning research 1mo ago Optimization of time-consuming experimental conditions using pseudo-experimental data guided by adaptive polynomial regression arXiv:2607.22238v1 Announce Type: new Abstract: Bayesian optimization (BO) is an optimization method that sequentially proposes the next candidate explainable variables for optimizing target variables by balancing exploration and exploitation. BO is often used under a limited… 27 arXiv — Machine Learning research 1mo ago LunarFM: A Shared Multimodal Representation of the Moon's Surface arXiv:2607.22408v1 Announce Type: new Abstract: The renewed global focus on lunar exploration, driven by the prospect of in-situ resource utilization and a sustained human presence on the Moon, has created growing demand for accurate, large-scale characterization of the lunar… 25 arXiv — Machine Learning research 1mo ago Hyperball May Not Be a Free Lunch arXiv:2607.22444v1 Announce Type: new Abstract: For scale-invariant deep networks, Hyperball-style optimizers have shown strong performance in large-scale training by fixing the norms of matrix-valued parameters and normalizing updates. However, the source of their advantage… 20 arXiv — Machine Learning research 1mo ago Susceptible Reservoir Architectures for Regime-Conditional Volatility Forecasting arXiv:2607.22491v1 Announce Type: new Abstract: Volatility forecasting is dominated by persistence and measurement noise, leaving limited residual structure for nonlinear models to exploit. We introduce Susceptible Architectures (SUSA), a reservoir-design principle for… 35 arXiv — Machine Learning research 1mo ago Prior laundering: learned priors with inherited, undetectable overconfidence arXiv:2607.21721v1 Announce Type: cross Abstract: Learned generative priors are increasingly used for ill-posed Bayesian inverse problems, their posterior uncertainty treated as earned from data. But training one requires truths, scarce in seismic and medical imaging, so the… 27 arXiv — NLP / Computation & Language research 1mo ago Khondo: A Multimodal Benchmark for Document Packet Splitting of Bangla Forms arXiv:2607.21780v1 Announce Type: new Abstract: Document packets, multiple documents concatenated into a single file, are common in government and administrative workflows, yet splitting them into their constituent documents is difficult, especially for low-resource languages.… 31 arXiv — NLP / Computation & Language research 1mo ago Biomedical Machine Translation for Low-Resource Arabic-Script Languages via Cross-Lingual Transfer and LoRA Adapter Merging arXiv:2607.22300v1 Announce Type: new Abstract: We present a systematic study of healthcare-domain cross-lingual transfer to address the scarcity of biomedical NMT resources for Arabic-script languages. We use Arabic and Persian as higher-resource pivots to improve translation… 37 arXiv — NLP / Computation & Language research 1mo ago A Factorial Study of Synthetic Data Generation for Low-Resource Machine Translation using Grammar Books arXiv:2607.22376v1 Announce Type: new Abstract: Most endangered languages lack the parallel data required for machine translation, despite the existence of descriptive grammar books. We introduce a pipeline that uses large language models to extract grammatical rules, example… 6 Vercel — AI dev-tools 1mo ago DeepsecBench: evaluating model performance in finding cybersecurity vulnerabilities Last week, OpenAI evaluated two models on an exploit benchmark within an isolated sandbox. Guardrails were reduced for testing, and the models found a vulnerability in their environment, accessed the internet, and reached Hugging Face's production database. No human directed the… 30 Vercel — AI dev-tools 1mo ago Nuxt July 2026 security advisory The Nuxt team has released Nuxt 4.5.1 and 3.21.10, along with @nuxt/devtools 3.3.1, to address eight security advisories, including a high-severity server-side remote code execution vulnerability. Vulnerabilities addressed Vulnerability Severity Advisory Server-side remote code… 30 r/LocalLLaMA community 1mo ago [OSS] Use case only possible with local inference at its core: an on-device LLM understands your entire life, then proactively offers to get your work done through computer use! Open-source & free :D Hey r/LocalLLaMA ! :D I wanna share a really cool fully OSS thing I've been building that's only possible with local models: truly proactive AI! All your existing LLM systems waits for a prompt. Truly proactive AI has to read your entire life, every single day (every file,… 19 r/LocalLLaMA community 1mo ago Sources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI   submitted by   /u/pscoutou [link]   [comments] 15 Page 9 of 10 · 500 articles ← Newer Older →