AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
Safety

Smart moves: Building resilient transportation systems with Google AI

DGX agent

What does transportation mean to you? For some, it’s making sure the train is on schedule so they can get to work on time. Maybe it’s making sure you have time connecting between flights. Maybe it’s a

safetygoogle-cloud-ai
13 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

so we built psql_bm25s. exact BM25 retrieval. native Postgres access method. ~23x faster than pg_search on the standard benchmark. retrieval…

DGX agent

so we built psql_bm25s. exact BM25 retrieval. native Postgres access method. ~23x faster than pg_search on the standard benchmark. retrieval stops being a budget item. the harness stops rationing. the

model-releasesemad-mostaque--x
13 May 2026
Safety

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning

DGX agent

arXiv:2605.12236v1 Announce Type: cross Abstract: Fine-tuning pre-trained robot policies with reinforcement learning (RL) often inherits the bottlenecks introduced by pre-training with behavioral clon

safetyarxiv-cs-lg
13 May 2026
Safety

Transformer-Based Autonomous Driving Models and Deployment-Oriented Compression: A Survey

DGX agent

arXiv:2304.10891v2 Announce Type: replace-cross Abstract: Transformer-based models are becoming a central paradigm in autonomous driving because they can capture long-range spatial dependencies, multi

safetyarxiv-cs-cv
13 May 2026
Research

Welfare as a Guiding Principle for Machine Learning -- From Compass, to Lens, to Roadmap

DGX agent

arXiv:2502.11981v3 Announce Type: replace Abstract: Decades of research in machine learning have given us powerful tools for making accurate predictions. But when used in social settings and on human

researcharxiv-cs-lg
13 May 2026
Industry

ZipRecruiter: IT and CS job postings are up 14.2% YoY in April; entry-level roles fell from 8.1% a year ago to 7.4%, while senior roles grew from 38.8% to 43.1% (Katherine Bindley/Wall Street Journal)

DGX agent

Katherine Bindley / Wall Street Journal: ZipRecruiter: IT and CS job postings are up 14.2% YoY in April; entry-level roles fell from 8.1% a year ago to 7.4%, while senior roles grew from 38.8% to 43.1

industrytechmeme
13 May 2026
Safety

A Cross-Layered Multi-Drone Coordination for Medical Supply Delivery during Disaster Response Management

DGX agent

arXiv:2605.09342v1 Announce Type: cross Abstract: Autonomous drone fleets have immense potential in medical supply delivery during disaster incident response. However, coordinating multiple drones in

safetyarxiv-cs-lg
12 May 2026
Model Releases

A Game Theoretic Free Energy Analysis of Higher Order Synergy in Attention Heads of Large Language Models

DGX agent

arXiv:2605.09515v1 Announce Type: new Abstract: Large language models rely on multihead attention, but interactions among heads remain poorly understood. We apply the Game Theoretic Free Energy Princi

model-releasesarxiv-cs-ai
12 May 2026
Safety

Anatomical Landmark-Guided Deep Reinforcement Learning for Autonomous Gastric Navigation

DGX agent

arXiv:2605.08269v1 Announce Type: new Abstract: Wireless capsule endoscopy (WCE) enables painless visualization of the gastrointestinal tract, but its diagnostic potential is limited by incomplete muc

safetyarxiv-cs-ro
12 May 2026
Research

Attribution-based Explanations for Markov Decision Processes

DGX agent

arXiv:2605.09780v1 Announce Type: new Abstract: Attribution techniques explain the outcome of an AI model by assigning a numerical score to its inputs. So far, these techniques have mainly focused on

researcharxiv-cs-ai
12 May 2026
Tutorials

Automate schema generation for intelligent document processing

DGX agent

In this post, we'll show you how our multi-document discovery feature solves this problem. It serves as an automated pre-processing step, analyzing unknown documents, clustering them by type, and gene

tutorialsaws-ml-blog
12 May 2026
Safety

Causal Explanations from the Geometric Properties of ReLU Neural Networks

DGX agent

arXiv:2605.10396v1 Announce Type: new Abstract: Neural networks have proved an effective means of learning control policies for autonomous systems, but these learned policies are difficult to understa

safetyarxiv-cs-lg
12 May 2026
Safety

Constraint-Aware Reinforcement Learning via Adaptive Action Scaling

DGX agent

arXiv:2510.11491v3 Announce Type: replace-cross Abstract: Safe reinforcement learning (RL) seeks to mitigate unsafe behaviors that arise from exploration during training by reducing constraint violati

safetyarxiv-cs-lg
12 May 2026
Tutorials

Cool idea from Nous Research. What if you could speed up long-context pretraining with a subquadratic wrapper that you remove before deploym…

DGX agent

Cool idea from Nous Research. What if you could speed up long-context pretraining with a subquadratic wrapper that you remove before deployment? That is the idea behind Lighthouse Attention. The metho

tutorialsdair-ai--x
12 May 2026
Research

Decentralized Conformal Novelty Detection via Quantized Model Exchange

DGX agent

arXiv:2605.08263v1 Announce Type: cross Abstract: This work studies decentralized novelty detection with global false discovery rate (FDR) control across heterogeneous composite null distributions, wi

researcharxiv-cs-lg
12 May 2026
Safety

Effective Explanations Support Planning Under Uncertainty

DGX agent

arXiv:2605.08406v1 Announce Type: cross Abstract: Explaining how to get from A to B can be challenging. It requires mentally simulating what the listener will do based on what they are told. To captur

safetyarxiv-cs-ai
12 May 2026
Local Ai

Efficient LLM Collaboration via Planning

DGX agent

arXiv:2506.11578v4 Announce Type: replace Abstract: Recently, large language models (LLMs) have demonstrated strong performance, ranging from simple to complex tasks. However, while large models achie

local-aiarxiv-cs-ai
12 May 2026
Safety

Generative Adversarial Post-Training Mitigates Reward Hacking in Live Human-AI Music Interaction

DGX agent

arXiv:2511.17879v4 Announce Type: replace Abstract: Most applications of generative AI involve a sequential interaction in which a person inputs a prompt and waits for a response, and where reaction t

safetyarxiv-cs-lg
12 May 2026
Hardware

How Imgix processes 8 billion images daily with G4 VMs powered by NVIDIA Blackwell

DGX agent

The modern web is extremely visual. People are busy and easily-distracted, and smart companies know they have just seconds to attract would-be customers with compelling images, videos, animations, and

hardwaregoogle-cloud-ai
12 May 2026
Safety

I sat down with @JonHernandezIA in Madrid to discuss the growing risks and impacts of AI and the urgent need to improve our social, politica…

DGX agent

I sat down with @JonHernandezIA in Madrid to discuss the growing risks and impacts of AI and the urgent need to improve our social, political, and technical safeguards. Thanks for an excellent convers

safetyyoshua-bengio--x
12 May 2026
Model Releases

Knowledge is Not Enough: Injecting RL Skills for Continual Adaptation

DGX agent

arXiv:2601.11258v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) face the 'knowledge cutoff' challenge, where their frozen parametric memory prevents direct internalization of ne

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

Large Language Models over Networks: Collaborative Intelligence under Resource Constraints

DGX agent

arXiv:2605.08626v1 Announce Type: cross Abstract: Large language models (LLMs) are transforming society, powering applications from smartphone assistants to autonomous driving. Yet cloud-based LLM ser

local-aiarxiv-cs-lg
12 May 2026
Safety

Learning to Compress Time-to-Control: A Reinforcement Learning Framework for Chronic Disease Management

DGX agent

arXiv:2605.09818v1 Announce Type: new Abstract: Reinforcement learning (RL) in healthcare has had mixed results, with reward sparsity, unreliable off-policy evaluation, and deployment-simulation gap a

safetyarxiv-cs-lg
12 May 2026
Hardware

Leveraging LLMs to Automate Energy-Aware Refactoring of Parallel Scientific Codes

DGX agent

arXiv:2505.02184v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used for generating parallel scientific codes, with a primary focus on generating functionally correct

hardwarearxiv-cs-ai
12 May 2026
Model Releases

LLM Wardens: Mitigating Adversarial Persuasion with Third-Party Conversational Oversight

DGX agent

arXiv:2605.08321v1 Announce Type: cross Abstract: LLMs are increasingly capable of persuasion, which raises the question of how to protect users against manipulation. In a preregistered user study (N=

model-releasesarxiv-cs-ai
12 May 2026
Hardware

LLMs for Secure Hardware Design and Related Problems: Opportunities and Challenges

DGX agent

arXiv:2605.10807v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into Electronic Design Automation (EDA) and hardware security is rapidly reshaping the semiconductor i

hardwarearxiv-cs-lg
12 May 2026
Safety

Long-Horizon Q-Learning: Accurate Value Learning via n-Step Inequalities

DGX agent

arXiv:2605.05812v2 Announce Type: replace Abstract: Off-policy, value-based reinforcement learning methods such as Q-learning are appealing because they can learn from arbitrary experience, including

safetyarxiv-cs-ai
12 May 2026
Safety

MARLaaS: Multi-Tenant Asynchronous Reinforcement Learning as a Service

DGX agent

arXiv:2605.08527v1 Announce Type: cross Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has significantly improved the reasoning capabilities of large language models (LLMs), particula

safetyarxiv-cs-ai
12 May 2026
Local Ai

Merlin: Deterministic Byte-Exact Deduplication for Lossless Context Optimization in Large Language Model Inference

DGX agent

arXiv:2605.09990v1 Announce Type: new Abstract: Data-intensive applications, ranging from large-scale retrieval systems to advanced data pipelines, are increasingly bottlenecked by the processing of h

local-aiarxiv-cs-cl
12 May 2026
Model Releases

Metal-Sci: A Scientific Compute Benchmark for Evolutionary LLM Kernel Search on Apple Silicon

DGX agent

arXiv:2605.09708v1 Announce Type: cross Abstract: We present Metal-Sci, a 10-task benchmark of scientific Apple Silicon Metal compute kernels spanning six optimization regimes (stencils, all-pairs in

model-releasesarxiv-cs-ai
12 May 2026
Safety

Multi-scale Predictive Representations for Goal-conditioned Reinforcement Learning

DGX agent

arXiv:2605.09364v1 Announce Type: new Abstract: This paper investigates robust representation learning in offline goal-conditioned reinforcement learning (GCRL). Particularly in sparse reward scenario

safetyarxiv-cs-lg
12 May 2026
Safety

MURPHY: Feedback-Aware GRPO with Retrospective Credit Assignment for Multi-Turn Code Generation

DGX agent

arXiv:2511.07833v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a standard recipe for post-training LLMs on reasoning tasks, with Group Relat

safetyarxiv-cs-ai
12 May 2026
Model Releases

Omni-DeepSearch: A Benchmark for Audio-Driven Omni-Modal Deep Search

DGX agent

arXiv:2605.08762v1 Announce Type: cross Abstract: Current omni-modal benchmarks mainly evaluate models under settings where multiple modalities are provided simultaneously, while the ability to start

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

OpenSGA: Efficient 3D Scene Graph Alignment in the Open World

DGX agent

arXiv:2605.10484v1 Announce Type: new Abstract: Scene graph alignment establishes object correspondences between two 3D scene graphs constructed from partially overlapping observations. This enables e

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

PaperFit: Vision-in-the-Loop Typesetting Optimization for Scientific Documents

DGX agent

arXiv:2605.10341v1 Announce Type: new Abstract: A LaTeX manuscript that compiles without error is not necessarily publication-ready. The resulting PDFs frequently suffer from misplaced floats, overflo

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Personal Visual Context Learning in Large Multimodal Models

DGX agent

arXiv:2605.10936v1 Announce Type: new Abstract: As wearable devices like smart glasses integrate Large Multimodal Models (LMMs) into the continuous first-person visual streams of individual users, the

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Pix2Fact: When Vision Is Not Enough -- Benchmarking Fine-Grained VQA with Web Verification on High-Resolution Real-World Scenes

DGX agent

arXiv:2602.00593v2 Announce Type: replace Abstract: Despite progress on general tasks, vision-language models (VLMs) still struggle with challenges that demand both fine-grained visual grounding and e

model-releasesarxiv-cs-cv
12 May 2026
Safety

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation

DGX agent

arXiv:2605.10118v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated exceptional general reasoning capabilities. However, their performance in embodied navigation remains hi

safetyarxiv-cs-ro
12 May 2026
Research

Position: Life-Logging Video Streams Make the Privacy-Utility Trade-off Inevitable

DGX agent

arXiv:2605.10404v1 Announce Type: new Abstract: With the growing prevalence of always-on hardware such as smart glasses, body cameras, and home security systems, life-logging visual sensing is becomin

researcharxiv-cs-cv
12 May 2026
Safety

Positive Alignment: Artificial Intelligence for Human Flourishing

DGX agent

arXiv:2605.10310v1 Announce Type: new Abstract: Existing alignment research is dominated by concerns about safety and preventing harm: safeguards, controllability, and compliance. This paradigm of ali

safetyarxiv-cs-ai
12 May 2026
Model Releases

PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

DGX agent

arXiv:2605.08687v1 Announce Type: cross Abstract: Data preparation is a central and time-consuming stage in data analysis workflows. Traditionally, commercial tools have relied on graphical user inter

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

PrimeKG-CL: A Continual Graph Learning Benchmark on Evolving Biomedical Knowledge Graphs

DGX agent

arXiv:2605.10529v1 Announce Type: new Abstract: Biomedical knowledge graphs underwrite drug repurposing and clinical decision support, yet the upstream ontologies they depend on update on independent

model-releasesarxiv-cs-ai
12 May 2026
Safety

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions

DGX agent

arXiv:2605.09893v1 Announce Type: cross Abstract: Large language models (LLMs) are often evaluated based on their stated values, yet these do not reliably translate into their actions, a discrepancy t

safetyarxiv-cs-ai
12 May 2026
Safety

Reinforcement Learning with Action Chunking

DGX agent

arXiv:2507.07969v4 Announce Type: replace-cross Abstract: We present Q-chunking, a simple yet effective recipe for improving reinforcement learning (RL) algorithms for long-horizon, sparse-reward task

safetyarxiv-cs-ai
12 May 2026
Safety

Revisiting Policy Gradients for Restricted Policy Classes: Escaping Myopic Local Optima with k-step Policy Gradients

DGX agent

arXiv:2605.10909v1 Announce Type: new Abstract: This work revisits standard policy gradient methods used on restricted policy classes, which are known to get stuck in suboptimal critical points. We id

safetyarxiv-cs-lg
12 May 2026
Model Releases

Scam2Prompt: A Scalable Framework for Auditing Malicious Scam Endpoints in Production LLMs

DGX agent

arXiv:2509.02372v3 Announce Type: replace-cross Abstract: Large Language Models have become critical to modern software development, but their reliance on uncurated web-scale datasets for training int

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SleepWalk: A Three-Tier Benchmark for Stress-Testing Instruction-Guided Vision-Language Navigation

DGX agent

arXiv:2605.10376v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have advanced rapidly in multimodal perception and language understanding, yet it remains unclear whether they can reliabl

model-releasesarxiv-cs-cv
12 May 2026
Industry

Sources: some Amazon employees are using in-house OpenClaw-like tool MeshClaw for unnecessary tasks to inflate AI token use after Amazon set weekly AI targets (Financial Times)

DGX agent

Financial Times: Sources: some Amazon employees are using in-house OpenClaw-like tool MeshClaw for unnecessary tasks to inflate AI token use after Amazon set weekly AI targets — In-house MeshClaw tool

industrytechmeme
12 May 2026
← Previous
1…349350351352353…367
Next →