AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
13 May 2026

Smart moves: Building resilient transportation systems with Google AI

SafetyDGX agent

What does transportation mean to you? For some, it’s making sure the train is on schedule so they can get to work on time. Maybe it’s making sure you have time connecting between flights. Maybe it’s a

so we built psql_bm25s. exact BM25 retrieval. native Postgres access method. ~23x faster than pg_search on the standard benchmark. retrieval…

Model ReleasesDGX agent

so we built psql_bm25s. exact BM25 retrieval. native Postgres access method. ~23x faster than pg_search on the standard benchmark. retrieval stops being a budget item. the harness stops rationing. the

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.12236v1 Announce Type: cross Abstract: Fine-tuning pre-trained robot policies with reinforcement learning (RL) often inherits the bottlenecks introduced by pre-training with behavioral clon

Transformer-Based Autonomous Driving Models and Deployment-Oriented Compression: A Survey

SafetyDGX agent

arXiv:2304.10891v2 Announce Type: replace-cross Abstract: Transformer-based models are becoming a central paradigm in autonomous driving because they can capture long-range spatial dependencies, multi

Welfare as a Guiding Principle for Machine Learning -- From Compass, to Lens, to Roadmap

ResearchDGX agent

arXiv:2502.11981v3 Announce Type: replace Abstract: Decades of research in machine learning have given us powerful tools for making accurate predictions. But when used in social settings and on human

ZipRecruiter: IT and CS job postings are up 14.2% YoY in April; entry-level roles fell from 8.1% a year ago to 7.4%, while senior roles grew from 38.8% to 43.1% (Katherine Bindley/Wall Street Journal)

IndustryDGX agent

Katherine Bindley / Wall Street Journal: ZipRecruiter: IT and CS job postings are up 14.2% YoY in April; entry-level roles fell from 8.1% a year ago to 7.4%, while senior roles grew from 38.8% to 43.1

12 May 2026

A Cross-Layered Multi-Drone Coordination for Medical Supply Delivery during Disaster Response Management

SafetyDGX agent

arXiv:2605.09342v1 Announce Type: cross Abstract: Autonomous drone fleets have immense potential in medical supply delivery during disaster incident response. However, coordinating multiple drones in

A Game Theoretic Free Energy Analysis of Higher Order Synergy in Attention Heads of Large Language Models

Model ReleasesDGX agent

arXiv:2605.09515v1 Announce Type: new Abstract: Large language models rely on multihead attention, but interactions among heads remain poorly understood. We apply the Game Theoretic Free Energy Princi

Anatomical Landmark-Guided Deep Reinforcement Learning for Autonomous Gastric Navigation

SafetyDGX agent

arXiv:2605.08269v1 Announce Type: new Abstract: Wireless capsule endoscopy (WCE) enables painless visualization of the gastrointestinal tract, but its diagnostic potential is limited by incomplete muc

Attribution-based Explanations for Markov Decision Processes

ResearchDGX agent

arXiv:2605.09780v1 Announce Type: new Abstract: Attribution techniques explain the outcome of an AI model by assigning a numerical score to its inputs. So far, these techniques have mainly focused on

Automate schema generation for intelligent document processing

TutorialsDGX agent

In this post, we'll show you how our multi-document discovery feature solves this problem. It serves as an automated pre-processing step, analyzing unknown documents, clustering them by type, and gene

Causal Explanations from the Geometric Properties of ReLU Neural Networks

SafetyDGX agent

arXiv:2605.10396v1 Announce Type: new Abstract: Neural networks have proved an effective means of learning control policies for autonomous systems, but these learned policies are difficult to understa

Constraint-Aware Reinforcement Learning via Adaptive Action Scaling

SafetyDGX agent

arXiv:2510.11491v3 Announce Type: replace-cross Abstract: Safe reinforcement learning (RL) seeks to mitigate unsafe behaviors that arise from exploration during training by reducing constraint violati

Cool idea from Nous Research. What if you could speed up long-context pretraining with a subquadratic wrapper that you remove before deploym…

TutorialsDGX agent

Cool idea from Nous Research. What if you could speed up long-context pretraining with a subquadratic wrapper that you remove before deployment? That is the idea behind Lighthouse Attention. The metho

Decentralized Conformal Novelty Detection via Quantized Model Exchange

ResearchDGX agent

arXiv:2605.08263v1 Announce Type: cross Abstract: This work studies decentralized novelty detection with global false discovery rate (FDR) control across heterogeneous composite null distributions, wi

Effective Explanations Support Planning Under Uncertainty

SafetyDGX agent

arXiv:2605.08406v1 Announce Type: cross Abstract: Explaining how to get from A to B can be challenging. It requires mentally simulating what the listener will do based on what they are told. To captur

Efficient LLM Collaboration via Planning

Local AiDGX agent

arXiv:2506.11578v4 Announce Type: replace Abstract: Recently, large language models (LLMs) have demonstrated strong performance, ranging from simple to complex tasks. However, while large models achie

Generative Adversarial Post-Training Mitigates Reward Hacking in Live Human-AI Music Interaction

SafetyDGX agent

arXiv:2511.17879v4 Announce Type: replace Abstract: Most applications of generative AI involve a sequential interaction in which a person inputs a prompt and waits for a response, and where reaction t

How Imgix processes 8 billion images daily with G4 VMs powered by NVIDIA Blackwell

HardwareDGX agent

The modern web is extremely visual. People are busy and easily-distracted, and smart companies know they have just seconds to attract would-be customers with compelling images, videos, animations, and

I sat down with @JonHernandezIA in Madrid to discuss the growing risks and impacts of AI and the urgent need to improve our social, politica…

SafetyDGX agent

I sat down with @JonHernandezIA in Madrid to discuss the growing risks and impacts of AI and the urgent need to improve our social, political, and technical safeguards. Thanks for an excellent convers

Knowledge is Not Enough: Injecting RL Skills for Continual Adaptation

Model ReleasesDGX agent

arXiv:2601.11258v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) face the 'knowledge cutoff' challenge, where their frozen parametric memory prevents direct internalization of ne

Large Language Models over Networks: Collaborative Intelligence under Resource Constraints

Local AiDGX agent

arXiv:2605.08626v1 Announce Type: cross Abstract: Large language models (LLMs) are transforming society, powering applications from smartphone assistants to autonomous driving. Yet cloud-based LLM ser

Learning to Compress Time-to-Control: A Reinforcement Learning Framework for Chronic Disease Management

SafetyDGX agent

arXiv:2605.09818v1 Announce Type: new Abstract: Reinforcement learning (RL) in healthcare has had mixed results, with reward sparsity, unreliable off-policy evaluation, and deployment-simulation gap a

Leveraging LLMs to Automate Energy-Aware Refactoring of Parallel Scientific Codes

HardwareDGX agent

arXiv:2505.02184v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used for generating parallel scientific codes, with a primary focus on generating functionally correct

LLM Wardens: Mitigating Adversarial Persuasion with Third-Party Conversational Oversight

Model ReleasesDGX agent

arXiv:2605.08321v1 Announce Type: cross Abstract: LLMs are increasingly capable of persuasion, which raises the question of how to protect users against manipulation. In a preregistered user study (N=

LLMs for Secure Hardware Design and Related Problems: Opportunities and Challenges

HardwareDGX agent

arXiv:2605.10807v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into Electronic Design Automation (EDA) and hardware security is rapidly reshaping the semiconductor i

Long-Horizon Q-Learning: Accurate Value Learning via n-Step Inequalities

SafetyDGX agent

arXiv:2605.05812v2 Announce Type: replace Abstract: Off-policy, value-based reinforcement learning methods such as Q-learning are appealing because they can learn from arbitrary experience, including

MARLaaS: Multi-Tenant Asynchronous Reinforcement Learning as a Service

SafetyDGX agent

arXiv:2605.08527v1 Announce Type: cross Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has significantly improved the reasoning capabilities of large language models (LLMs), particula

Merlin: Deterministic Byte-Exact Deduplication for Lossless Context Optimization in Large Language Model Inference

Local AiDGX agent

arXiv:2605.09990v1 Announce Type: new Abstract: Data-intensive applications, ranging from large-scale retrieval systems to advanced data pipelines, are increasingly bottlenecked by the processing of h

Metal-Sci: A Scientific Compute Benchmark for Evolutionary LLM Kernel Search on Apple Silicon

Model ReleasesDGX agent

arXiv:2605.09708v1 Announce Type: cross Abstract: We present Metal-Sci, a 10-task benchmark of scientific Apple Silicon Metal compute kernels spanning six optimization regimes (stencils, all-pairs in

Multi-scale Predictive Representations for Goal-conditioned Reinforcement Learning

SafetyDGX agent

arXiv:2605.09364v1 Announce Type: new Abstract: This paper investigates robust representation learning in offline goal-conditioned reinforcement learning (GCRL). Particularly in sparse reward scenario

MURPHY: Feedback-Aware GRPO with Retrospective Credit Assignment for Multi-Turn Code Generation

SafetyDGX agent

arXiv:2511.07833v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a standard recipe for post-training LLMs on reasoning tasks, with Group Relat

Omni-DeepSearch: A Benchmark for Audio-Driven Omni-Modal Deep Search

Model ReleasesDGX agent

arXiv:2605.08762v1 Announce Type: cross Abstract: Current omni-modal benchmarks mainly evaluate models under settings where multiple modalities are provided simultaneously, while the ability to start

OpenSGA: Efficient 3D Scene Graph Alignment in the Open World

Model ReleasesDGX agent

arXiv:2605.10484v1 Announce Type: new Abstract: Scene graph alignment establishes object correspondences between two 3D scene graphs constructed from partially overlapping observations. This enables e

PaperFit: Vision-in-the-Loop Typesetting Optimization for Scientific Documents

Model ReleasesDGX agent

arXiv:2605.10341v1 Announce Type: new Abstract: A LaTeX manuscript that compiles without error is not necessarily publication-ready. The resulting PDFs frequently suffer from misplaced floats, overflo

Personal Visual Context Learning in Large Multimodal Models

Model ReleasesDGX agent

arXiv:2605.10936v1 Announce Type: new Abstract: As wearable devices like smart glasses integrate Large Multimodal Models (LMMs) into the continuous first-person visual streams of individual users, the

Pix2Fact: When Vision Is Not Enough -- Benchmarking Fine-Grained VQA with Web Verification on High-Resolution Real-World Scenes

Model ReleasesDGX agent

arXiv:2602.00593v2 Announce Type: replace Abstract: Despite progress on general tasks, vision-language models (VLMs) still struggle with challenges that demand both fine-grained visual grounding and e

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation

SafetyDGX agent

arXiv:2605.10118v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated exceptional general reasoning capabilities. However, their performance in embodied navigation remains hi

Position: Life-Logging Video Streams Make the Privacy-Utility Trade-off Inevitable

ResearchDGX agent

arXiv:2605.10404v1 Announce Type: new Abstract: With the growing prevalence of always-on hardware such as smart glasses, body cameras, and home security systems, life-logging visual sensing is becomin

Positive Alignment: Artificial Intelligence for Human Flourishing

SafetyDGX agent

arXiv:2605.10310v1 Announce Type: new Abstract: Existing alignment research is dominated by concerns about safety and preventing harm: safeguards, controllability, and compliance. This paradigm of ali

PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

Model ReleasesDGX agent

arXiv:2605.08687v1 Announce Type: cross Abstract: Data preparation is a central and time-consuming stage in data analysis workflows. Traditionally, commercial tools have relied on graphical user inter

PrimeKG-CL: A Continual Graph Learning Benchmark on Evolving Biomedical Knowledge Graphs

Model ReleasesDGX agent

arXiv:2605.10529v1 Announce Type: new Abstract: Biomedical knowledge graphs underwrite drug repurposing and clinical decision support, yet the upstream ontologies they depend on update on independent

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions

SafetyDGX agent

arXiv:2605.09893v1 Announce Type: cross Abstract: Large language models (LLMs) are often evaluated based on their stated values, yet these do not reliably translate into their actions, a discrepancy t

Reinforcement Learning with Action Chunking

SafetyDGX agent

arXiv:2507.07969v4 Announce Type: replace-cross Abstract: We present Q-chunking, a simple yet effective recipe for improving reinforcement learning (RL) algorithms for long-horizon, sparse-reward task

Revisiting Policy Gradients for Restricted Policy Classes: Escaping Myopic Local Optima with k-step Policy Gradients

SafetyDGX agent

arXiv:2605.10909v1 Announce Type: new Abstract: This work revisits standard policy gradient methods used on restricted policy classes, which are known to get stuck in suboptimal critical points. We id

Scam2Prompt: A Scalable Framework for Auditing Malicious Scam Endpoints in Production LLMs

Model ReleasesDGX agent

arXiv:2509.02372v3 Announce Type: replace-cross Abstract: Large Language Models have become critical to modern software development, but their reliance on uncurated web-scale datasets for training int

SleepWalk: A Three-Tier Benchmark for Stress-Testing Instruction-Guided Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2605.10376v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have advanced rapidly in multimodal perception and language understanding, yet it remains unclear whether they can reliabl

Sources: some Amazon employees are using in-house OpenClaw-like tool MeshClaw for unnecessary tasks to inflate AI token use after Amazon set weekly AI targets (Financial Times)

IndustryDGX agent

Financial Times: Sources: some Amazon employees are using in-house OpenClaw-like tool MeshClaw for unnecessary tasks to inflate AI token use after Amazon set weekly AI targets — In-house MeshClaw tool

TeleResilienceBench: Quantifying Resilience for LLM Reasoning in Telecommunications

Model ReleasesDGX agent

arXiv:2605.09929v1 Announce Type: new Abstract: Deploying large language models in telecommunications requires more than task accuracy. In realistic workflows, a model may inherit partially completed

There will be no AI jobpocalypse. The story that AI will lead to massive unemployment is stoking unnecessary fear. AI — like any other techn…

SafetyDGX agent

There will be no AI jobpocalypse. The story that AI will lead to massive unemployment is stoking unnecessary fear. AI — like any other technology — does affect jobs, but telling overblown stories of l

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍

SafetyDGX agent

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍 If anyone builds it, everyone thrives. Over the past decade, a lot of important work on AI alignment has focus

Towards Customized Multimodal Role-Play

SafetyDGX agent

arXiv:2605.08129v1 Announce Type: new Abstract: Unified multimodal understanding and generation models enable richer human-AI interaction. Yet jointly customizing a character's persona, dialogue style

Training-Free Cultural Alignment of Large Language Models via Persona Disagreement

SafetyDGX agent

arXiv:2605.10843v1 Announce Type: cross Abstract: Large language models increasingly mediate decisions that turn on moral judgement, yet a growing body of evidence shows that their implicit preference

UserGPT Technical Report

Model ReleasesDGX agent

arXiv:2605.08766v1 Announce Type: cross Abstract: Personalized user understanding from large-scale digital traces remains a fundamental challenge. Traditional user profiling methods rely on discrimina

V-ABS: Action-Observer Driven Beam Search for Dynamic Visual Reasoning

SafetyDGX agent

arXiv:2605.10172v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have achieved remarkable success in general perception, yet complex multi-step visual reasoning remains a per

We just crossed 1,000,000 public datasets on Hugging Face! That's petabytes of data available that millions of AI builders are downloading, …

IndustryDGX agent

We just crossed 1,000,000 public datasets on Hugging Face! That's petabytes of data available that millions of AI builders are downloading, analyzing, and training AI models on every day! What's inter

When a Robot is More Capable than a Human: Learning from Constrained Demonstrators

SafetyDGX agent

arXiv:2510.09096v3 Announce Type: replace-cross Abstract: Learning from demonstrations enables experts to teach robots complex tasks using interfaces such as kinesthetic teaching, joystick control, an

11 May 2026

2.5-D Decomposition for LLM-Based Spatial Construction

Model ReleasesDGX agent

arXiv:2605.07066v1 Announce Type: new Abstract: Autonomous systems that build structures from natural-language instructions need reliable spatial reasoning, yet large language models (LLMs) make syste

Ask Patients with Patience: Enabling LLMs for Human-Centric Medical Dialogue with Grounded Reasoning

Model ReleasesDGX agent

arXiv:2502.07143v3 Announce Type: replace Abstract: The severe shortage of medical doctors limits access to timely and reliable healthcare, leaving millions underserved. Large language models (LLMs) o

CktFormalizer: Autoformalization of Natural Language into Circuit Representations

HardwareDGX agent

arXiv:2605.07782v1 Announce Type: new Abstract: LLMs can generate hardware descriptions from natural language specifications, but the resulting Verilog often contains width mismatches, combinational l

← Previous
1…279280281282283…294
Next →