News / #security Tag Security 500 articles archived under #security · RSS Sign in to follow Hugging Face Daily Papers research 4d ago CyberFactory: Scaling Cyber Security Capabilities with Instances from the Wild Abstract CyberFactory is an open-source framework that builds agentic training data from real vulnerabilities to train Aegis, improving open-weight cybersecurity performance across proof-of-concept generation, patching, and question answering. Generated by… 13 r/MachineLearning community 4d ago Millwright — experimenting with an end-to-end machine learning framework in Rust [P] I've been working on an open-source project called Millwright , an attempt to explore what an end-to-end machine learning workflow could look like in Rust. https://millwright-rs.dev/ This started while I was learning and building ML tooling in Rust. I kept finding capable… 16 Hugging Face Daily Papers research 4d ago Game2World Engine: Unlocking In-the-Wild Gameplay Videos for World Model Training Abstract A framework for removing gameplay UI from videos enables cleaner training data for video world models, improving reward metrics and outperforms mask-based removal methods. Generated by thinkingmachines/Inkling-Small Video games provide a scalable source of training data… 11 r/LocalLLaMA community 4d ago Open Source Kernel in Qwen3.6-35B-A3B for AMD MI350X: 78,498 output tok/s on 8 GPUs So here's the thing, almost everyone use NVIDIA to run their LLMs, we also do the same, a lot of people we've met use like RTX PRO 6000 or even H100, B300 It seems like everyone eyes is looking at NVIDIA. However we do the math that the raw power alone on AMD GPU MI350X is… 9 arXiv — Machine Learning research 4d ago UHI-Bench: Benchmarking Dual-Source Urban Heat Island Modeling Across Cities in Diverse Climate Regimes arXiv:2608.23857v1 Announce Type: new Abstract: Urban heat islands (UHIs) are intensifying under climate change, exacerbating thermal exposure risks. Their two primary observations, land surface temperature UHI (LST-UHI) and near-surface air temperature UHI (AirT-UHI), capture… 21 arXiv — Machine Learning research 4d ago Physics-Integrated Operator Learning via Gaussian Splatting Representations arXiv:2608.24049v1 Announce Type: new Abstract: Neural operators provide efficient surrogates for spatiotemporal PDE systems, but purely data-driven formulations often accumulate substantial errors during long-horizon autoregressive prediction and may fail to exploit available… 38 arXiv — Machine Learning research 4d ago Joint Distribution Alignment for Universal Domain Adaptation arXiv:2608.24429v1 Announce Type: new Abstract: Unsupervised domain adaptation (UDA) has been widely concerned in the fields of machine learning, pattern recognition, and computer vision. Traditional UDA learning usually assumes that the label spaces of the source and target… 17 arXiv — Machine Learning research 4d ago WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation arXiv:2608.24479v1 Announce Type: new Abstract: Massively parallel simulation changes the data regime in which off-policy reinforcement learning (RL) is trained, challenging stabilizers designed for data-limited replay. Through controlled experiments across eight benchmark… 35 arXiv — Machine Learning research 4d ago Data Leakage Inflates Generalizability of Power Outage Prediction Models arXiv:2608.24665v1 Announce Type: new Abstract: Power outage prediction models are increasingly used in assessments of climate-driven infrastructure risk, yet current evaluation practices obscure whether these models generalize to the novel conditions such applications require.… 6 arXiv — Machine Learning research 4d ago When May an Agent Stop? Evidence-Carrying Termination for Tool-Using LLMs arXiv:2608.23623v1 Announce Type: cross Abstract: Tool-using agents must decide when to stop. Existing systems already gate terminal success, certify execution traces, or enforce runtime polici es, but do not test this particular receipt-, scope-, and closed-replay design at the… 8 Hugging Face Daily Papers research 4d ago WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation Abstract Off-policy reinforcement learning stabilizers vary with data availability, motivating regime-aware algorithms that adapt normalization and Q-function clipping to improve efficiency across CPU and GPU-parallel training. Generated by thinkingmachines/Inkling-Small… 19 The Information — AI news-outlet 4d ago Exclusive: Second OpenAI Sales Exec to Return to Salesforce OpenAI recruited heavily from Salesforce to sell its AI to big businesses. Now, some sales people are returning to Salesforce, deepening the shake-up in OpenAI’s enterprise sales team. Peter Doolan, the former chief customer officer of Slack who joined OpenAI in April, has… 19 Dwarkesh Podcast news-outlet 4d ago Dylan Patel – Anthropic & OpenAI will have most of the world’s compute by 2028 "Every force is screeching towards centralization." 33 arXiv — NLP / Computation & Language research 5d ago KSE-Web: An Analysis of Hybrid Retrieval and LLM-Assisted Query Expansion for Low-Resource Khmer Semantic Search arXiv:2608.21365v1 Announce Type: new Abstract: As a low-resource language, Khmer presents several retrieval challenges, including limited annotated data, ambiguous word boundaries, weak support in multilingual embedding models, and frequent mixed Khmer-English usage. This paper… 31 arXiv — NLP / Computation & Language research 5d ago Mitigating Database Leakage in RAG Systems with Keyword-Grounded Fact Substitution arXiv:2608.21656v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has emerged as a powerful paradigm for combining large language models (LLMs) with external knowledge sources. However, RAG systems remain vulnerable to prompt injection attacks, which may… 16 arXiv — NLP / Computation & Language research 5d ago L\"etzCross: A Cross-Lingual Page-Level Benchmark for Multimodal Retrieval over Luxembourgish Documents arXiv:2608.21714v1 Announce Type: new Abstract: Recent page-image retrievers such as ColPali have improved retrieval over visually rich documents, yet little is known about how they behave in cross-lingual, low-resource settings. We introduce L\"etzCross, a benchmark for… 9 arXiv — NLP / Computation & Language research 5d ago Improving Few-Step Language Flows with Untied Self-Conditioning arXiv:2608.22244v1 Announce Type: new Abstract: Flow-matching language models refine all token positions in parallel and can trade sampling steps for latency, yet generation quality still degrades sharply with few sampling steps. We trace a source of this degradation to a… 12 arXiv — NLP / Computation & Language research 5d ago Length-Adaptive Decoding for Masked Diffusion Machine Translation arXiv:2608.22274v1 Announce Type: new Abstract: Machine translation tests masked diffusion language models (dLLMs) because every source token must be rendered faithfully, while fixed canvas decoding must choose target length before denoising. Existing masked diffusion decoding… 5 Hugging Face Daily Papers research 5d ago Prime Agent: A Self-Improving RLM Harness Abstract Prime Agent is an open-source harness that uses recursive subagents, persistent computation, and agent-to-agent coordination to extend language models' long-horizon capabilities across coding and reasoning tasks. Generated by thinkingmachines/Inkling-Small Language… 22 Hugging Face Daily Papers research 5d ago TLive-Omni: An Omni-Modal Understanding Model for E-Commerce Live Streaming Abstract TLive-Omni is an omni-modal model for live-commerce that unifies image, video, audio, and text via timestamped token grouping, staged supervised training, and reinforcement fine-tuning with verifiable feedback to enable accurate real-time understanding. Generated by… 24 The Information — AI news-outlet 5d ago Exclusive: Hugging Face Annualized Revenue Jumps 50% to $150 Million AI startup Hugging Face, known for its repository of open-source models, is generating more than $150 million in annualized revenue, a 50% increase from two months ago , according to a person with direct knowledge of the matter. The decade-old startup is nearing a deal to sell… 13 r/LocalLLaMA community 5d ago [2608.16157] FreeToken: Efficient Edge-Native MoE Serving with Bandwidth-Adaptive Execution Source of Claims: https://x.com/Andy_ShuoYang/status/2090856976880472439 Your gaming PC can now serve frontier models at interactive speed using official checkpoints without extreme quantization! Qwen3.6 35B → 8GB RTX 4060 laptop @ 39 tok/s DeepSeek-V4-Flash 284B → RTX 5090… 22 llama.cpp releases dev-tools 5d ago b10614 metal: per-op source split + parallel compile ( #26561 ) metal : per-op source split + parallel compile ( #24021 ) preliminary extract common header op source split split metallib into 8 libs && load in parallel derive kernel->library routing from functionNames x-macro lib list… 21 Smol AI News news-outlet 6d ago not much happened today **Agent harnesses** are becoming a key optimization focus, with NVIDIA research showing traditional skill checks poorly predict agent usefulness and proposing a new metric called **"Skill Lift"**. Open-source implementations of **persistent and self-modifying agents** like… 27 Smol AI News news-outlet 6d ago not much happened today **Microduck**, a **25 cm open-source biped robot** from **Pollen Robotics** and **Hugging Face**, priced at **$399** and shipping before Christmas, features **15 actuators** and a rich sensor suite including camera, LiDAR, NFC, Bluetooth, and Wi-Fi. It supports… 32 arXiv — Machine Learning research 6d ago When Clean Data Hurts: Learning with Monotone Corruptions Beyond Binary Classification arXiv:2608.20480v1 Announce Type: new Abstract: Optimal learners are tailored to exploit the i.i.d.\ data assumption underlying the classic PAC model. What if an i.i.d.\ training sample were corrupted with correctly labeled examples drawn from an otherwise unrelated, even… 4 arXiv — Machine Learning research 6d ago Faults That Fortify: CNN Adversarial Robustness via GPU Undervolting arXiv:2608.20572v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) face a dual challenge: vulnerability to adversarial attacks and prohibitive training cost. Adversarial training is effective but expensive, a burden that grows as learning shifts to the… 15 arXiv — Machine Learning research 6d ago RiskTraf: Risk-Extrapolated Residual Learning for Multi-Variate Traffic Flow Prediction arXiv:2608.20656v1 Announce Type: new Abstract: Traffic sensors commonly record flow, speed, and occupancy, but standard traffic flow forecasting benchmarks and models rarely exploit all three raw measurements reliably. Although speed and occupancy provide sensor-native… 28 arXiv — Machine Learning research 6d ago Trojaning the Alignment: Stealthy Backdoor Attacks against Graph Foundation Models arXiv:2608.20991v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) on text-attributed graphs (TAGs) align graph representations with language semantics to support transferable graph learning. Despite these advantages, the backdoor vulnerability of GFMs on TAGs… 20 arXiv — Machine Learning research 6d ago Capturing Cardiac Cyclicity through Phase-Equivariant Self-Supervised Learning arXiv:2608.21147v1 Announce Type: new Abstract: The cyclic structure of physiological processes offers a natural prior for self-supervised representation learning, and the cardiac cycle provides a particularly well-defined setting in which to exploit it. We derive a… 24 arXiv — Machine Learning research 6d ago Rigorous Evaluation of Large Language Models for Malaria Drug Discovery: Trade-offs in Performance, Scale, and Resource Utility arXiv:2608.20418v1 Announce Type: cross Abstract: We introduce Malaria-Instruct, a curated instruction-following dataset derived from the ChEMBL Legacy Malaria corpus for Malaria virtual screening, and conduct a systematic evaluation of five open-source LLMs; Gemma-2 2B/9B,… 36 arXiv — Machine Learning research 6d ago Keyed Provenance Watermarking with Complementary Lattice-Based Secure Aggregation for Federated Learning arXiv:2608.20580v1 Announce Type: cross Abstract: Federated learning (FL) is vulnerable to multi-level attacks. However, existing methods address them separately, leaving FL exposed to data leakage, unauthorized reuse, and malicious gradient manipulation. In this work, we… 26 arXiv — Machine Learning research 6d ago Predicting Resource Efficient Hamiltonian Decomposition for Continuous-Time Quantum Walk Simulations arXiv:2608.20660v1 Announce Type: cross Abstract: Simulating a continuous-time quantum walk (CTQW) on a graph in the circuit model of quantum computing requires decomposing its Hamiltonian into terms that can be Trotterized into hardware-native gates. We consider two such… 15 arXiv — Machine Learning research 6d ago AudioWorldSim: Realistic Binaural Audio Datasets For World Models arXiv:2608.21075v1 Announce Type: cross Abstract: This technical report presents AudioWorldSim, an open-source platform designed to generate realistic binaural audio datasets and advance research in audio-based machine learning, particularly world models. Built as a custom… 37 arXiv — NLP / Computation & Language research 6d ago Building and Evaluating a Synthetic Bengali Speech Resource for Telecom Customer Care arXiv:2608.20346v1 Announce Type: new Abstract: Speech systems used in customer-facing applications often require domain-specific language coverage. We present a synthetic Bengali speech dataset for telecom customer-care scenarios. The dataset contains 10,000 audio-text pairs,… 8 arXiv — NLP / Computation & Language research 6d ago AsmEvo: Agentic Assembly-Level Optimization of AMD GPU Kernels with Functional Equivalence Verification arXiv:2608.20711v1 Announce Type: new Abstract: High-performance ML systems increasingly rely on GPU kernels whose editable source is unavailable, generated, or too distant from final machine code to expose remaining optimizations. Existing LLM kernel optimizers and autotuners… 31 arXiv — NLP / Computation & Language research 6d ago Ontology-Driven Structural Regularization for Document-Level Relation Extraction arXiv:2608.20856v1 Announce Type: new Abstract: Document-Level Relation Extraction (DocRE) relies heavily on costly manually annotated datasets, while large distant supervision resources such as DocRED distant remain underexploited due to noise. We show that a critical yet… 12 Hugging Face Daily Papers research 6d ago Towards Faithful Simulation of Human Shopping Behavior Abstract RecVerse is a GUI-grounded agent that uses hierarchical memory and trajectory-level reinforcement learning to simulate realistic multi-turn e-commerce shopping sessions. Generated by thinkingmachines/Inkling-Small Simulating realistic user shopping behavior underpins… 23 r/LocalLLaMA community 6d ago Any upcoming models to be excited about? I'm kind of new to the community and this sub is my only source of information, so I'd thought of asking if there are any upcoming models you guys are looking forward to.   submitted by   /u/zippydazoop [link]   [comments] 14 The Information — AI news-outlet 6d ago Nvidia, Salesforce Are in Spotlight This Week Some of the hardest-working people in tech right now have to be the corporate communications folks at Nvidia, which is pretty much never out of the news. Aside from updates on its latest AI chips, Nvidia seems to be investing in almost every part of the AI sector, from data… 26 r/LocalLLaMA community 7d ago Closed AI has been real quiet since Qwen 3.8 27B dropped. This is something I've noticed. Back when GLM 5.2 and Kimi K3 launched, there was a media push on pushing how dangerous open source models are. We know why they were doing this; these open models are good enough to devalue paid closed models. I think the fear mongering was to… 16 r/MachineLearning community 7d ago I built an open-source roguelike specifically for training game-playing agents [P] Hey everyone! I wanted to share something I’ve been working on. I was inspired by projects from DeepMind and OpenAI, but noticed that most games are prohibitively difficult to integrate with an agent harness. So I built DelveRL from the ground up as a human-playable game with a… 36 r/LocalLLaMA community 8d ago Llama.cpp version 0.2.0 is out! You can find the changelog and source code here: https://github.com/ggml-org/llama.cpp/releases/tag/v0.2.0 Associated pre-build is here: https://github.com/ggml-org/llama.cpp/releases/tag/b10566   submitted by   /u/PhilippeEiffel [link]   [comments] 32 r/LocalLLaMA community 8d ago Bro wtf, Qwen Lab cooked with Qwen 3.8 27B, it's so fucking good Context: Earlier, open-source large models like Mimo V2.5 Pro, DeepSeek V4 Pro (first version), and Kimi K2.5 used to struggle with this prompt, and Qwen 3.6 27B couldn't even render the globe properly. But now, Qwen 3.8 27B is so much better than the previous model. I hope qwen… 29 llama.cpp releases dev-tools 8d ago b10567 ci : run ccache-clear as the last step of release jobs ( #27503 ) ci : run ccache-clear as the last step of release jobs Assisted-by: pi:llama.cpp/Qwen3.8-27B update disabled job too to force rebase Co-authored-by: Sigbjørn Skjæret sigbjorn.skjaeret@huggingface.co Website:… 11 r/MachineLearning community 8d ago repo2nb 0.2.0, convert a GitHub repo into a Kaggle/Colab notebook (dependency resolution, reverse mode, incremental sync) [P] repo2nb is an open-source CLI that converts a GitHub repo into a runnable Kaggle or Colab notebook: walks the file tree, resolves dependencies, and generates cells, instead of you doing that by hand for a repo you didn't write (a paper's code, a tutorial, someone else's… 21 Hugging Face Daily Papers research 8d ago Thinking in a Low-Resource Language: What SFT Builds, What RL Fixes, What Accuracy Cannot See Abstract Fine-tuning large mixture-of-experts models on a low-resource language shifts reasoning into that language without harming accuracy, while reinforcement learning with verifiable rewards fixes formatting and leakage defects. Generated by thinkingmachines/Inkling-Small… 11 arXiv — Machine Learning research 9d ago Quantifying Event Impacts on Time Series via Multiscale Contrastive Learning arXiv:2608.19447v1 Announce Type: new Abstract: Shocks that spread through the web, such as cybersecurity breach disclosures, can abruptly disrupt financial time series and cause substantial abnormal losses. While these events are disclosed as discrete records through news… 31 arXiv — Machine Learning research 9d ago Reliable Neural Collapse Approximation for Open-World Test-Time Adaptation arXiv:2608.19890v1 Announce Type: new Abstract: Test-Time Adaptation (TTA) methods aim to bridge the domain gap between the source and target domains. However, traditional TTA methods become ineffective when the label distribution shift occurs, a challenge commonly referred to… 6 arXiv — Machine Learning research 9d ago Multi-Source Wasserstein Distributionally Robust Graph Learning arXiv:2608.19914v1 Announce Type: new Abstract: Network topology inference from graph signals is central to graph signal processing with applications in neuroscience, sensor, and social networks. In practice, target-domain samples are scarce while heterogeneous source-domain… 32 Page 2 of 10 · 500 articles ← Newer Older →