AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,237 results
23 Jun 2026

Open models, global networks: How AT&T and GSMA are accelerating telecom innovation with Gemma

Model ReleasesDGX agent

Telecommunications is an incredibly complex, highly specialized domain. Modern mobile networks are inherently multi-vendor, featuring diverse and often proprietary data structures. While AI has made m

R2HandoverSim: A Simulation Framework and Benchmark for Robot-to-Human Object Handovers

Model ReleasesDGX agent

arXiv:2606.21011v1 Announce Type: new Abstract: We present R2HandoverSim, a simulation benchmark for robot-to-human (R2H) object handovers. Although R2H handover methods have advanced rapidly, the lac

SparseWorld: Enhancing End-to-End Autonomous Driving via World Models with Sparse Scene Representation

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.24354v2 Announce Type: replace Abstract: Recently, world models have made significant progress in enhancing end-to-end driving systems through both future situation forecasting and improved

Zero-Shot Vision-Language Models for Classroom Engagement Recognition: A Benchmark Study of Prompt Sensitivity and Cross-Dataset Generalization

Model ReleasesDGX agent

arXiv:2606.21861v1 Announce Type: new Abstract: Automated classroom engagement recognition holds substantial promise for scalable learning analytics, yet the suitability of modern Vision-Language Mode

22 Jun 2026

1,250 hp hybrid Corvette shatters the Pikes Peak production record

ApplicationsDGX agent

Chevrolet's Corvette ZR1X, powered by a hybrid all-wheel drive system with 1,250 horsepower, set a new Pikes Peak production car record at the 104th running of the hill climb. Driven by IndyCar vetera

Daybreak: Tools for securing every organization in the world

Model ReleasesDGX agent

Daybreak is an OpenAI initiative focused on developing and distributing security tools designed to protect organizations globally from cyber threats. The program aims to democratize access to advanced

GM installs robots at flagship EV factory after laying off 1,300 workers

IndustryDGX agent

General Motors temporarily laid off 1,300 workers at its Factory Zero EV plant due to sluggish demand for battery-powered vehicles , following the discontinuation of the $7,500 federal EV tax credit .

How Anthropic may have talked itself into an AI export ban

IndustryDGX agent

The US government ordered Anthropic to cut off foreign access to its Mythos 5 and Claude Fable 5 AI models in June 2026, citing national security concerns related to jailbreak vulnerabilities that cou

Red-Teaming after Mythos — Zico Kolter & Matt Fredrikson, Gray Swan

ToolsDGX agent

This episode likely discusses red-teaming methodologies and adversarial testing approaches in AI systems, featuring security researchers Zico Kolter and Matt Fredrikson from Carnegie Mellon University

11 Jun 2026

Agent Skill Evaluation and Evolution: Frameworks and Benchmarks

Model ReleasesDGX agent

arXiv:2606.11435v1 Announce Type: new Abstract: The growth of agent skills has transformed how agentic systems are built, evaluated, and deployed. As skill libraries continue to scale, rigorous evalua

Making Models Unmergeable via Scaling-Sensitive Loss Landscape

Model ReleasesDGX agent

arXiv:2601.21898v2 Announce Type: replace Abstract: The rise of model hubs has made it easier to access reusable model components, making model merging a practical tool for combining capabilities. Yet

10 Jun 2026

LLM-Based Code Documentation Generation and Multi-Judge Evaluation

Model ReleasesDGX agent

arXiv:2606.09852v1 Announce Type: cross Abstract: High-quality source code documentation is vital yet often neglected, especially in critical domains like healthcare where reliability and maintainabil

Uncertainty-Aware Motion Planning for Autonomous Driving in Mixed Traffic Environment

Model ReleasesDGX agent

arXiv:2606.09958v1 Announce Type: cross Abstract: In mixed-traffic environments where autonomous and human-driven vehicles may co-exist, motion planning for autonomous vehicles requires anticipating t

9 Jun 2026

AI-Native Closed-Loop Security for 6G-Enabled Cyber-Physical Systems: From Edge Detection to Network-Wide Mitigation

Local AiDGX agent

arXiv:2606.08173v1 Announce Type: cross Abstract: In sixth-generation (6G) networks, billions of cyber-physical systems (CPSs) - autonomous vehicles, smart grids, industrial robots, and remote-surgica

Anthropic says Fable 5 has invisible safeguards that use prompt modification, steering vectors, or PEFT to limit its effectiveness for building frontier LLMs (Matthias Bastian/The Decoder)

Model ReleasesDGX agent

Matthias Bastian / The Decoder: Anthropic says Fable 5 has invisible safeguards that use prompt modification, steering vectors, or PEFT to limit its effectiveness for building frontier LLMs — Key Poin

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.07805v1 Announce Type: new Abstract: The rapid evolution of Large Language Models (LLMs) from passive assistants to autonomous, execution-capable agents has introduced critical operational

Explaining Black-Box Language Models: Learning to Optimize Linguistically-Structured Word Subsets

Model ReleasesDGX agent

arXiv:2606.08497v1 Announce Type: new Abstract: As deep language models (DLMs) are increasingly deployed in high-stakes domains such as healthcare, understanding their decision rationale becomes param

Initial impressions of Claude Fable 5

Model ReleasesDGX agent

I didn't have early access to today's Claude Fable 5 release, but I've spent the past ~5.5 hours putting it through its paces. My initial impressions are that this is something of a beast. It's slow,

Learning from Human Driving: A Human-in-the-Loop Online Behavior Cloning Framework for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.08170v1 Announce Type: new Abstract: With the evolution of large foundation models (LFMs), data-driven autonomous driving has made significant strides. However, existing paradigms still fac

PLAGUE: Plug-and-play framework for Lifelong Adaptive Generation of Multi-turn Exploits

Model ReleasesDGX agent

arXiv:2510.17947v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are improving at an exceptional rate. With the advent of agentic workflows, multi-turn dialogue has become the de

ProbeAct: Probe-Guided Training-Free Failure Recovery in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2606.09740v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models demonstrate strong perfor-1 mance on language-conditioned robotic manipulation within their training dis-2 tribution

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks

Model ReleasesDGX agent

arXiv:2606.07968v1 Announce Type: cross Abstract: Reasoning-capable large language models can be induced to spend their generation budget on injected decoy tasks rather than answering the user's quest

RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations

Model ReleasesDGX agent

arXiv:2606.08376v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across socially consequential domains, reports of AI-related harms and failures have

Safe Polytope-in-Polytope Motion Planning and Control with Control Barrier Functions

Local AiDGX agent

arXiv:2606.09719v1 Announce Type: new Abstract: Autonomous mobile robots operating in tight environments require motion planning frameworks that account for the physical footprint of the robot. Simpli

Signals Are Not States: Neuro-Symbolic Safeguards for Culturally Aware Classroom AI

Model ReleasesDGX agent

arXiv:2603.22793v2 Announce Type: replace Abstract: Classroom AI systems increasingly infer high-level educational states such as engagement, confusion, collaboration, participation, and instructional

Sound and Complete Neurosymbolic Reasoning with LLM-Grounded Interpretations

Local AiDGX agent

arXiv:2507.09751v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated impressive capabilities in natural language understanding and generation, but exhibit problems with l

Strained Coherence: A Pre-Failure Signal in Coding Agent Execution Trajectories

Model ReleasesDGX agent

arXiv:2606.07889v1 Announce Type: cross Abstract: LLM-based coding agents sometimes acknowledge a problem in their own reasoning and then proceed anyway. We call this pattern strained coherence: a saf

Structural Decoupling: A Scaffold-Flow Theory of Generalization and Alignment

Local AiDGX agent

arXiv:2506.20699v2 Announce Type: replace Abstract: Learning in non-stationary and multi-context environments requires more than ordinary within-task generalization. A system must also discover which

TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking

Model ReleasesDGX agent

arXiv:2602.03224v2 Announce Type: replace Abstract: Test-time evolution of agent memory represents a pivotal paradigm for advancing AGI, as it strengthens complex reasoning through experience accumula

They didn’t mean pause AI research, they meant pause *your* AI research

TutorialsDGX agent

Jeremy Howard argues that calls to pause AI research are selectively applied, with restrictions primarily targeting independent researchers while well-resourced labs continue development, creating an

VATS: Exploiting Implicit Authority in Error-Path Injection via Systematic Mutation

Model ReleasesDGX agent

arXiv:2606.07992v1 Announce Type: new Abstract: As the Model Context Protocol (MCP) standardizes tool-calling for autonomous agents, it introduces a critical, unexamined attack surface: the error-hand

What EU regulations does to AI

HardwareDGX agent

EU regulations, particularly the AI Act, establish comprehensive compliance requirements for AI systems including risk-based classification, transparency obligations, and restrictions on high-risk app

8 Jun 2026

Endogenous Resistance to Activation Steering in Language Models

Model ReleasesDGX agent

arXiv:2602.06941v2 Announce Type: replace-cross Abstract: Large language models can recover mid-generation from task-misaligned activation steering, producing explicit verbal restarts (e.g., ``wait, t

Hierarchical Certified Semantic Commitment for Byzantine-Resilient LLM-Agent Collaboration

Model ReleasesDGX agent

arXiv:2606.07316v1 Announce Type: cross Abstract: Byzantine collaboration among large-language-model agents requires a finality-control primitive: given delivered stochastic, structured natural-langua

LLM-Guided Evolution for Medical Decision Pipelines

Model ReleasesDGX agent

arXiv:2606.07342v1 Announce Type: new Abstract: Adapting large language models (LLMs) to clinical workflows often requires costly fine-tuning or manual prompt and pipeline engineering. We study LLM-gu

Multi-Objective Preference Optimization: Improving Human Alignment of Generative Models

Model ReleasesDGX agent

arXiv:2505.10892v2 Announce Type: replace Abstract: Post-training LLMs with RLHF and preference optimization methods (e.g., DPO, IPO) has greatly improved alignment, yet these approaches assume a sing

REMEDI: A Benchmark for Retention and Unlearning Evaluation in Multi-label Clinical Disease Inference

Model ReleasesDGX agent

arXiv:2606.07141v1 Announce Type: cross Abstract: Language models trained for clinical disease inference are trained on patient data, which may include sensitive and private information, and data owne

When Large Language Models Fail in Healthcare: Evaluating Sensitivity to Prompt Variations

Model ReleasesDGX agent

arXiv:2606.07237v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in healthcare for tasks such as clinical question answering, diagnosis support, and report summariz

6 Jun 2026

Evaluating Agentic Configuration Repair for Computer Networks

Model ReleasesDGX agent

arXiv:2606.06212v1 Announce Type: new Abstract: Misconfigurations in computer networks remain a major source of critical Internet outages. Research is turning to Large Language Models (LLMs) to automa

Trust, but Don't Verify: Epistemic Blind Spots in LLM Source Evaluation

Model ReleasesDGX agent

arXiv:2606.05403v1 Announce Type: cross Abstract: Language models increasingly act as epistemic proxies, synthesizing evidence from multiple sources to inform decisions. Whether they evaluate the qual

5 Jun 2026

CLASH: Evaluating Language Models on Judging High-Stakes Dilemmas from Multiple Perspectives

Model ReleasesDGX agent

arXiv:2504.10823v4 Announce Type: replace Abstract: Navigating dilemmas involving conflicting values is challenging even for humans in high-stakes domains, let alone for AI, yet prior work has been li

CLEAR: Cognition and Latent Evaluation for Adaptive Routing in End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.06219v1 Announce Type: new Abstract: End-to-end autonomous driving models often struggle to balance multi-modal maneuver generation with real-time inference constraints. While diffusion mod

Seeing is Believing? Evaluating Vision-Language Model Susceptibility in Agent-to-Agent Multimodal Persuasion

Model ReleasesDGX agent

arXiv:2510.22768v2 Announce Type: replace Abstract: As autonomous agents increasingly interact, they inevitably attempt to influence one another. While prior work in text-only settings has explored th

Unlocking dependable responses with Gemini Enterprise Agent Platform’s Agentic RAG

Model ReleasesDGX agent

Google's RAG Engine securely connects private enterprise data to LLMs to improve answer accuracy and reduce hallucinations , making it a key component of the Gemini Enterprise Agent Platform for build

VASO: Formally Verifiable Self-Evolving Skills for Physical AI Agents

Local AiDGX agent

arXiv:2606.05395v1 Announce Type: new Abstract: Reusable robot skills are becoming the basic units through which embodied agents turn open-ended instructions into long-horizon physical behavior. We ar

4 Jun 2026

An Open-Source Two-Stage Computer Vision Pipeline for Fine-Grained Vehicle Classification using Vision Transformers

Model ReleasesDGX agent

arXiv:2606.05149v1 Announce Type: new Abstract: Vehicle body type is a significant determinant of cyclist injury severity in overtaking crashes, yet automated tools for classifying vehicles into injur

Andon Labs' Real-World AI Evals: Claude calls the FBI, AI CEOs, price cartels, Butter-Bench, & Luna https://latent.space/p/andon @andonlabs …

Model ReleasesDGX agent

Andon Labs' Real-World AI Evals: Claude calls the FBI, AI CEOs, price cartels, Butter-Bench, & Luna https://latent.space/p/andon @andonlabs cofounders @lukaspet and @axelbacklund explain why dollar-de

Long Live Fine-Tuning: Task-Specific Transformers Outperform Zero-Shot LLMs for Misinformation Response Classification on Reddit

Model ReleasesDGX agent

arXiv:2606.04274v1 Announce Type: new Abstract: As large language models (LLMs) become default tools for online information verification, an implicit assumption follows them: that scale and general ca

MedForge: Interpretable Medical Deepfake Detection via Forgery-aware Reasoning

Model ReleasesDGX agent

arXiv:2603.18577v2 Announce Type: replace Abstract: Text-guided image editors can now manipulate authentic medical scans with high fidelity, enabling lesion implantation/removal that threatens clinica

The Canadian AI strategy unveiled today advocates for the development of technology that is safe, ethical, trustworthy, and that benefits so…

Model ReleasesDGX agent

The Canadian AI strategy unveiled today advocates for the development of technology that is safe, ethical, trustworthy, and that benefits society as a whole—these are exactly the principles that need

Toward Pre-Deployment Assurance for Enterprise AI Agents: Ontology-Grounded Simulation and Trust Certification

Model ReleasesDGX agent

arXiv:2606.04037v1 Announce Type: new Abstract: Pre-deployment verification of enterprise artificial intelligence (AI) agents remains a critical gap between large language model (LLM) capability bench

3 Jun 2026

Acceptance-Test-Driven Evaluation Protocols for Business-Centric LLM Systems

Model ReleasesDGX agent

arXiv:2606.02755v1 Announce Type: cross Abstract: Large language model (LLM) applications are increasingly expected to satisfy deterministic institutional requirements while relying on probabilistic g

Agent libOS: A Library-OS-Inspired Runtime for Long-Running, Capability-Controlled LLM Agents

Local AiDGX agent

arXiv:2606.03895v1 Announce Type: cross Abstract: Large language model (LLM) agents are evolving from request-response assistants into long-running software actors: they maintain state across model ca

How Quantization Changes Interpretable Features: A Sparse Autoencoder Analysis of Language Models

Model ReleasesDGX agent

arXiv:2606.03002v1 Announce Type: cross Abstract: Quantization is a standard path to deploying large language models, and a quantized model is typically judged acceptable when its perplexity or downst

LAMP: Data-Efficient Linear Affine Weight-Space Models for Parameter-Controlled 3D Shape Generation and Extrapolation

Model ReleasesDGX agent

arXiv:2510.22491v3 Announce Type: replace-cross Abstract: Generating high-fidelity 3D geometries under explicit parameter constraints is central to engineering design, yet current methods often requir

Reliability-Guided Depth Fusion for Glare-Resilient Navigation Costmaps

Model ReleasesDGX agent

arXiv:2606.03421v1 Announce Type: new Abstract: Specular glare on reflective floors, glass boundaries, and glossy indoor surfaces frequently corrupts active-stereo RGB-D depth measurements, producing

Toward a Modular Architecture for Embedded AI Agent Systems at the Edge

Local AiDGX agent

arXiv:2606.02862v1 Announce Type: new Abstract: The rise of Large Language Models (LLMs) has enabled agentic AI capable of complex reasoning and tool use; however, deploying such autonomy in pervasive

TriEval: A Resource-Efficient Pipeline for LLM Bias, Toxicity, and Truthfulness Assessment

Model ReleasesDGX agent

arXiv:2606.03036v1 Announce Type: new Abstract: LLMs have evolved from basic chatbots to the backbone of the AI ecosystem, now widely used in healthcare, schools, and government services. The domain-w

VidMsg: A Benchmark for Implicit Message Inference in Short Videos

Model ReleasesDGX agent

arXiv:2606.03635v1 Announce Type: cross Abstract: Understanding short online videos involves more than identifying visible objects and actions; video makers often include an underlying message or purp

What's the most unhinged thing you've used an uncensored Ollama model for? Also... what are the best uncensored models right now?

Local AiDGX agent

This Reddit discussion from r/ollama explores user experiences with uncensored Ollama language models, featuring anecdotal accounts of unusual or extreme use cases and recommendations for popular unce

← Previous
1…229230231232233…238
Next →