Introducing CodeMender: an AI agent for code security
Introducing CodeMender: an AI agent for code security Using advanced AI to fix critical software vulnerabilities
313 stories from 26 sources
Introducing CodeMender: an AI agent for code security Using advanced AI to fix critical software vulnerabilities
Some Large Language Models Exhibit Consistent Risk Attitudes arXiv:2607.16197v1 Announce Type: new Abstract: As artificial intelligence systems are deployed in open-ended, high-stakes settings, a critical dimension rema…
A Survey on the Verification of Reinforcement Learning Policies arXiv:2607.16210v1 Announce Type: new Abstract: Reinforcement learning (RL) is increasingly applied in complex, safety-critical domains, yet the lack of ri…
Bridging the Information Gap: Semantic Densification and Hindsight Distillation for Cold-Start Prediction arXiv:2607.17070v1 Announce Type: new Abstract: New-user cold-start is a critical bottleneck for e-commerce platf…
Benchmarking Machine Learning Models for Multi-Omics-Based Breast Cancer Prediction arXiv:2607.16250v1 Announce Type: cross Abstract: Estrogen Receptor (ER) status is a critical biomarker in breast cancer diagnosis, pro…
Detection, Attribution, Narration: An End-to-End Pipeline for Explainable Money Mule Identification arXiv:2607.17586v1 Announce Type: new Abstract: Money mule accounts are critical facilitators of financial fraud, yet d…
LLM Unlearning for Cyber Defense: A Survey on Methods, Challenges, and Emerging Threats arXiv:2607.16227v1 Announce Type: cross Abstract: LLMs are increasingly deployed in security-critical systems across healthcare, fi…
The 2026 Agent Confidence Index: Where 300 builders see real momentum At Microsoft, building trustworthy AI agents is as critical as building powerful ones. New research from the 2026 Agent Confidence Index shows where…
Deepening our partnership with the UK AI Security Institute Google DeepMind and UK AI Security Institute (AISI) strengthen collaboration on critical AI safety and security research
FakeGit Campaign Uses 7,600 GitHub Repositories to Spread SmartLoader Malware Cybersecurity researchers have discovered nearly 7,600 malicious GitHub repositories, out of which more than 800 pose as artificial intellige…
The Real AI Threat Is Blind Trust AI models left to both interpret and execute commands eliminate critical cybersecurity oversight.
Google Launches Gemini 3.5 Flash Cyber AI to Find and Fix Software Vulnerabilities Google's DeepMind on Tuesday announced the release of Gemini 3.5 Flash Cyber, a specialized artificial intelligence (AI) model built ato…
SlotGuard: Stop Oversharing Private Local Context in LLM Agent Transcri arXiv:2607.17147v1 Announce Type: new Abstract: LLM agents can leak privacy (e.g., paths, emails) and credentials (e.g., API keys) as agent observa…
Do Speech Tokens Leak Voiceprints? Speaker Inversion Attacks Against End-to-End Speech Language Models arXiv:2607.16870v1 Announce Type: cross Abstract: End-to-end speech language models increasingly represent user spee…
Is "Knowing It's Malicious Enough?" Evaluating LLMs for Fine-Grained Malware Behavior Auditing arXiv:2509.14335v2 Announce Type: replace Abstract: Automated malware classifiers achieve strong detection performance, but…
TRACE: Trajectory-Based Safety Patch Learning for LLM Post-Training Realignment arXiv:2607.16242v1 Announce Type: cross Abstract: Fine-Tuning-as-a-Service (FTaaS) platforms let users train large language models (LLMs) o…
New AI Model Shows How to Evacuate for Fires One Safe Step at a Time A NIST-led team has created a new AI model that can identify safe evacuation routes in a single-story floor plan during a fire, with a multilevel vers…
CAISI Issues Request for Information About Securing AI Agent Systems The Center for AI Standards and Innovation (CAISI) at the U.S. Department of Commerce’s National Institute of Standards and Technology (NIST) has publ…
CAISI Evaluation of DeepSeek AI Models Finds Shortcomings and Risks The Center for AI Standards and Innovation at NIST evaluated several leading models from DeepSeek, an AI company based in the People’s Republic of Chin…
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? arXiv:2605.26548v2 Announce Type: replace Abstract: Finding a real vulnerability in complicated systems is a challenging, long-horizon task…
US threatens sanctions against Chinese AI models over IP theft Treasury Secretary Scott Bessent said the U.S. could sanction Chinese open AI models over alleged IP theft, expanding the Trump administration's campaign to…
The Shift: A New Era of AI Regulation Explore how recent US export controls on frontier AI models like Anthropic's Fable signal a new era of regulatory uncertainty. Learn how security leaders can build resilient AI stra…
Open-Source Android AI Agents Could Let Invisible Screen Text Run Code on Host PCs An Android app that can draw over other windows and write to shared storage can slip instructions to the AI agent driving that phone, in…
Russian-Speaking Hacker Uses Google Gemini CLI to Control Botnet of Eight Dental Clinic PCs A solo Russian-speaking threat actor known as "bandcampro" outsourced a chunk of their operations to Google's open-source Gemin…
New Agent Data Injection Attack Can Make AI Agents Misclick or Run Attacker Commands Ask an AI agent to summarize the reviews on a product page, and a single planted review can make it click "Buy Now" instead. Ask a cod…
OpenAI’s GPT-Red Automates Prompt Injection Testing to Harden GPT-5.6 Sol OpenAI has disclosed details of GPT-Red, an internal automated red-teaming model that scales prompt injection vulnerability discovery with an aim…
TuxBot v3 Evolution Shows Signs of LLM-Assisted IoT Botnet Development Cybersecurity researchers have disclosed details of a previously unreported Internet-of-Things (IoT) botnet framework dubbed TuxBot v3 Evolution tha…
Cisco Launches Low-Cost AI Models for Source Code Security The open-weight Antares models are designed to pinpoint known vulnerabilities in codebases faster and at a fraction of the cost of larger AI models. The post Ci…
Google Is Building an AI Chip Just for Gemini—And Investors Already Moved On It Codenamed Frozen v2, Google's new server chip would bake part of Gemini's architecture directly into hardware for a projected 6–10x efficie…
Anthropic’s $1.5 billion book piracy settlement approved by judge A federal judge has signed off on Anthropic's $1.5 billion class action settlement with authors who accused the company of training its AI models on copy…
How to build interactive experiences with canvases Canvases turn AI into interactive workspaces where you can visualize information, explore workflows, and take action across complex tasks. The post How to build interac…
The Download: Chinese AI divides the White House, and a record copyright payout This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. C…
Advancing next-gen AI with materials science innovation The conversation about AI often centers on algorithms, computing power, or huge investments in new semiconductor fabrication plants and hyperscale data centers. Bu…
Rater State Bias in RLHF Preference Data: An Audit Framework arXiv:2607.16195v1 Announce Type: new Abstract: We identify a structured confound in Reinforcement Learning from Human Feedback (RLHF). Pairwise preference la…
Design and Validation of a Lightweight 1D CNN for Affective Touch Classification in Soft Plush Companions arXiv:2607.16196v1 Announce Type: new Abstract: Soft, sensorized companions offer a physically safe and emotional…
A Survey on GNN-based Link Prediction: Techniques, Applications, and Challenges arXiv:2607.16198v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have emerged as the leading paradigm for link prediction, enab…
PlanFlip: Attacking Multi-Agent LLM Systems via Planning-Phase Prompt Injection arXiv:2607.16199v1 Announce Type: new Abstract: Multi-agent LLM systems increasingly rely on a Planner to decompose goals into sub-task seq…
Democratizing AI with Small Language Models: Structured Benchmarking and Parameter-Efficient Fine-Tuning for Local Deployment arXiv:2607.16202v1 Announce Type: new Abstract: AI democratization is not primarily a questio…
It Takes 8 Tokens: Weak-to-Strong Off-Policy RL via Auxiliary Branches arXiv:2607.16205v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has emerged as a standard approach for enhancing reas…
PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Coverage Optimization arXiv:2607.16206v1 Announce Type: new Abstract: This paper introduces PPO-HSC (Proximal Policy Optimization with H…
JUMP: Single-Pass Membership Inference on Fine-Tuned Diffusion Language Models arXiv:2607.16207v1 Announce Type: new Abstract: Membership inference attacks (MIAs) test whether a candidate example appeared in a model's t…
Accurate and Efficient Long-Term Memory for LLM Agents arXiv:2607.16211v1 Announce Type: new Abstract: LLM agents augmented with persistent memory can recall past interactions, but existing systems suffer from two limit…
Symbolic Augmentation Closes a Canonical-Equivalence Blind Spot in Neural Fact-Checkers arXiv:2607.16212v1 Announce Type: new Abstract: Large language models hallucinate numbers and units when summarizing scientific tex…
SelKV: Selective KV Cache Merging with Per-Token Merge-or-Drop and Attention Compensation arXiv:2607.16213v1 Announce Type: new Abstract: Large Language Models (LLMs) generate text autoregressively, relying on a key-val…
RAIL Guard: Closing the Evaluation-to-Remediation Gap in Responsible AI for LLM Agents arXiv:2607.16215v1 Announce Type: new Abstract: Existing guardrail systems for large language model agents operate as binary classif…
LaCache: Exact Caching and Precision-Adaptive Inference for Diffusion Large Language Models arXiv:2607.16339v1 Announce Type: new Abstract: Diffusion-based Large Language Models(DLLMs) enable parallel generation via Sem…
Interactive Task Alignment as a POMDP arXiv:2607.16412v1 Announce Type: new Abstract: Current benchmarks for language models primarily evaluate execution on fully specified tasks. However, real user tasks are often ambi…
When to Plan: Learning to Select Between Reactive Control and Deliberative Planning arXiv:2607.16421v1 Announce Type: new Abstract: It has long been recognized that humans have the ability to switch between fast, reacti…
SEER: Supervised Learning to Control Energetic Reasoning arXiv:2607.16523v1 Announce Type: new Abstract: One of the main strengths of Constraint Programming is the ability to reduce the search space via propagation. How…
Nonuniformity Principle in Human-AI Coworking arXiv:2607.16530v1 Announce Type: new Abstract: As generative AI is increasingly applied to automate multi-step and high-stake workflows, human judgment and involvement rema…
FST.ai 2.5: Explainable and Uncertainty-Aware AI for Olympic and Para-Taekwondo Decision Support, Athlete Digital Twins, and Federation-Scale Analytics arXiv:2607.16597v1 Announce Type: new Abstract: The rapid digitalis…
Just A Rather Very Intelligent Spoken Agent arXiv:2607.16610v1 Announce Type: new Abstract: Long-horizon AI agents are becoming increasingly capable, yet their interaction with users remains surprisingly thin. In most w…
A Research Prototype for Closed-Loop Generative Design of Customized Foot Orthoses via Semantic-Physics Alignment arXiv:2607.16631v1 Announce Type: new Abstract: Translating unstructured clinical prescriptions into pati…
TopoTuner: Topological Finetuning of Large Language Models arXiv:2607.16637v1 Announce Type: new Abstract: Full fine-tuning remains a strong way to adapt pretrained LLMs, but it updates all weights and can be expensive.…
Diversity-Oriented Fine-Tuning for Uncertainty-Based Hallucination Detection arXiv:2607.16643v1 Announce Type: new Abstract: Existing hallucination detection methods are typically conducted at the inference stage, witho…
RECON: Benchmarking Agent Memory for Compositional Reasoning over Long Contexts arXiv:2607.16716v1 Announce Type: new Abstract: Large language models and LLM-based agents are widely used as personal chat assistants, ent…
Constraint-Anchored Reasoning Traces arXiv:2607.16727v1 Announce Type: new Abstract: Autoregressive multimodal large language models (MLLMs) suffer from error snowballing: a single incorrect inference early in a chainof…
FUSAR-R1: A Large-Scale Reasoning Model for Intelligent Interpretation of SAR Images arXiv:2607.16819v1 Announce Type: new Abstract: In recent years, large-scale vision-language models have been driving a paradigm shift…
From Overload to Insights: How AI Agents Can Support Scientists in Analyzing Complex Data arXiv:2607.16845v1 Announce Type: new Abstract: Scientists at European XFEL conduct experiments that generate very large and comp…
AgentBrew: Lifelong Knowledge Brewing from Strong Teachers to Weak LLM Agents arXiv:2607.16851v1 Announce Type: new Abstract: Deploying LLM agents typically requires a compact test-time student, even if a stronger teach…