AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,762 results
25 May 2026

Disentangling Interaction and Bias Effects in Opinion Dynamics of Large Language Models

SafetyDGX agent

arXiv:2509.06858v2 Announce Type: replace-cross Abstract: Large Language Models are increasingly used to simulate human opinion dynamics, yet the effect of genuine interaction is often obscured by sys

MadEvolve: Evolutionary Optimization of Trading Systems with Large Language Models

Model ReleasesDGX agent

arXiv:2605.23007v1 Announce Type: cross Abstract: We explore the application of LLM-driven algorithm optimization to several common tasks in quantitative finance. MadEvolve, a general-purpose algorith

PIMbot: A Self-Adaptive Attack Framework for Adversarial Manipulation of Multi-Robot Reinforcement Learning

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.23027v1 Announce Type: new Abstract: Recent research has demonstrated the potential of reinforcement learning in effective multi-robot collaboration, particularly in social dilemmas where r

23 May 2026

Shipper’s “After Automation” is the best framing I’ve read of why AI keeps generating more work, not less. The line that stuck with me: the …

Model ReleasesDGX agent

Shipper’s “After Automation” is the best framing I’ve read of why AI keeps generating more work, not less. The line that stuck with me: the frame is not the framer. Models climb whatever benchmark we

22 May 2026

LACO: Adaptive Latent Communication for Collaborative Driving

SafetyDGX agent

arXiv:2605.22504v1 Announce Type: cross Abstract: Collaborative driving aims to improve safety and efficiency by enabling connected vehicles to coordinate under partial observability. Recent approache

Open-World Evaluations for Measuring Frontier AI Capabilities

Model ReleasesDGX agent

arXiv:2605.20520v1 Announce Type: new Abstract: Benchmark-based evaluation remains important for tracking frontier AI progress. But it can both overstate and understate deployed capability because it

21 May 2026

Frontier: Towards Comprehensive and Accurate LLM Inference Simulation

HardwareDGX agent

arXiv:2605.21312v1 Announce Type: cross Abstract: Modern LLM serving is no longer homogeneous or monolithic. Production systems now combine disaggregated execution, complex parallelism, runtime optimi

FWIW, I think this moves up my AI timelines a bit. I think the next milestone will be 'Artificial *Grothendieck* Intelligence' (AGrI): defin…

IndustryDGX agent

FWIW, I think this moves up my AI timelines a bit. I think the next milestone will be 'Artificial *Grothendieck* Intelligence' (AGrI): defining new general mathematical structures to solve the hardest

Hyper-V2X: Hypernetworks for Estimating Epistemic and Aleatoric Uncertainty in Cooperative Bird's-Eye-View Semantic Segmentation

Model ReleasesDGX agent

arXiv:2605.21309v1 Announce Type: new Abstract: Cooperative perception enabled by Vehicle-to-Everything (V2X) communication enhances autonomous driving safety by creating a unified environmental repre

It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs

SafetyDGX agent

arXiv:2605.20258v1 Announce Type: new Abstract: Contextual Integrity (CI) defines privacy not merely as keeping information hidden, but as governing information flows according to the norms of a given

Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX

SafetyDGX agent

arXiv:2605.20577v1 Announce Type: cross Abstract: Riichi Mahjong is a multi-player, imperfect-information game characterized by stochasticity and high-dimensional state spaces. These attributes presen

Mechanisms of Misgeneralization in Physical Sequence Modeling

Local AiDGX agent

arXiv:2605.20299v1 Announce Type: new Abstract: Generative sequence models are often trained to plan motion in physical domains, from robotics to mechanical simulations. When constructing a dataset to

Tutor-Student Reinforcement Learning: A Dynamic Curriculum for Robust Deepfake Detection

SafetyDGX agent

arXiv:2603.24139v2 Announce Type: replace Abstract: Standard supervised training for deepfake detection treats all samples with uniform importance, which can be suboptimal for learning robust and gene

20 May 2026

From Refusal to Recovery: A Control-Theoretic Approach to Generative AI Guardrails

SafetyDGX agent

arXiv:2510.13727v2 Announce Type: replace Abstract: Generative AI systems are increasingly assisting and acting on behalf of end users in practical settings, from digital shopping assistants to next-g

fun read on real tradeoffs & design decisions we debated when designing Engine for the data scale that customers produce one common thread i…

Model ReleasesDGX agent

fun read on real tradeoffs & design decisions we debated when designing Engine for the data scale that customers produce one common thread is that we’re pretty strong supporters of just giving the age

Guiding Neuro-Symbolic Scenario Generation with Spatio-Temporal Logic

SafetyDGX agent

arXiv:2605.19038v1 Announce Type: cross Abstract: The rapid advancement of autonomous driving (AD) technologies has outpaced the development of robust safety evaluation methods. Conventional testing r

When Critics Disagree: Adaptive Reward Poisoning Attacks in RIS-Aided Wireless Control System

SafetyDGX agent

arXiv:2605.20037v1 Announce Type: cross Abstract: Reward-poisoning attacks present a significant risk to learning-based wireless control systems. Given this, we propose a Disagreement-Guided Reward Po

19 May 2026

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback

Model ReleasesDGX agent

arXiv:2605.18073v1 Announce Type: cross Abstract: Large Language Models (LLMs) demonstrate strong potential for automated code generation, yet their ability to iteratively refine solutions using execu

BacktestBench: Benchmarking Large Language Models for Automated Quantitative Strategy Backtesting

Model ReleasesDGX agent

arXiv:2605.17937v1 Announce Type: cross Abstract: Quantitative backtesting is essential for evaluating trading strategies but remains hampered by high technical barriers and limited scalability. While

Evolve the Method, Not the Prompts: Evolutionary Synthesis of Jailbreak Attacks on LLMs

Model ReleasesDGX agent

arXiv:2511.12710v2 Announce Type: replace Abstract: Automated red teaming frameworks for Large Language Models (LLMs) have become increasingly sophisticated, yet many still formulate attack optimizati

Gemini 3.5 Flash: more expensive, but Google plan to use it for everything

Model ReleasesDGX agent

Today at Google I/O, Google released Gemini 3.5 Flash. This one skipped the -preview modifier and went straight to general availability, and Google appear to be using it for a whole lot of their key p

LEAF: A Living Benchmark for Event-Augmented Forecasting

Model ReleasesDGX agent

arXiv:2605.16358v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly applied to forecasting. To evaluate this capability while mitigating pre-training data contamination, se

LISTEN to Your Preferences: An LLM Framework for Multi-Objective Selection

SafetyDGX agent

arXiv:2510.25799v2 Announce Type: replace Abstract: Human experts often struggle to select the best option from a large set of items with multiple competing objectives, a process bottlenecked by the d

oldsymbol{f}-OPD: Stabilizing Long-Horizon On-Policy Distillation with Freshness-Aware Control

SafetyDGX agent

arXiv:2605.17862v1 Announce Type: cross Abstract: Scaling on-policy distillation (OPD) for large language models (LLMs) confronts a fundamental tension: asynchronous execution is necessary for system

QuantFPFlow: Quantum Amplitude Estimation for Fokker--Planck Policy Optimisation in Continuous Reinforcement Learning

SafetyDGX agent

arXiv:2605.16429v1 Announce Type: cross Abstract: We introduce extbf{QuantFPFlow}, a reinforcement learning framework that integrates quantum amplitude estimation into the Fokker--Planck~(FP) formulat

Shared Backbone PPO for Multi-UAV Communication Coverage with Connection Preservation

SafetyDGX agent

arXiv:2605.17999v1 Announce Type: new Abstract: This paper proposes a Shared Backbone Proximal Policy Optimization (Shared Backbone PPO) algorithm. By sharing the base module between the Actor and Cri

SocialMemBench: Are AI Memory Systems Ready for Social Group Settings?

Model ReleasesDGX agent

arXiv:2605.17789v1 Announce Type: cross Abstract: Memory systems for AI assistants were built for single-user dialogue and fail characteristically when applied to multi-party social group settings. Th

TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?

Model ReleasesDGX agent

arXiv:2605.18025v1 Announce Type: new Abstract: While Large Language Models have achieved remarkable integration in various vertical scenarios, their deployment in the telecommunications domain remain

18 May 2026

See Before You Code: Learning Visual Priors for Spatially Aware Educational Animation Generation

Local AiDGX agent

arXiv:2605.15585v1 Announce Type: new Abstract: Large language models can generate executable code for educational animations, but the resulting renders often exhibit visual defects, including element

17 May 2026

we do not post AIE videos with bullshit brainrot hype lingo, and this is the consequence: the entire AIE back catalog is being reposted by '…

Model ReleasesDGX agent

we do not post AIE videos with bullshit brainrot hype lingo, and this is the consequence: the entire AIE back catalog is being reposted by 'influence operators' almost daily, without credit to speaker

16 May 2026

Warelay -> OpenClaw

ToolsDGX agent

In preparation for a lightning talk I'm giving at PyCon US this afternoon I decided to figure out how many names OpenClaw has actually had since that first commit back in November. Thanks to this firs

15 May 2026

FrontierSmith: Synthesizing Open-Ended Coding Problems at Scale

ApplicationsDGX agent

arXiv:2605.14445v1 Announce Type: new Abstract: Many real-world coding challenges are open-ended and admit no known optimal solution. Yet, recent progress in LLM coding has focused on well-defined tas

14 May 2026

Adaptive Smooth Tchebycheff Attention for Multi-Objective Policy Optimization

SafetyDGX agent

arXiv:2605.12771v1 Announce Type: cross Abstract: Multi-objective reinforcement learning in robotic domains requires balancing complex, non-convex trade-offs between conflicting objectives. While line

Fireworks Training Platform continues to expand. Today GLM 5.1 LoRA RL is now live via Training API: SFT, DPO, and full RL on a 200K context…

Model ReleasesDGX agent

Fireworks Training Platform continues to expand. Today GLM 5.1 LoRA RL is now live via Training API: SFT, DPO, and full RL on a 200K context window → custom loss functions or smart defaults. No usage

Flow Matching for Offline Reinforcement Learning with Discrete Actions

SafetyDGX agent

arXiv:2602.06138v2 Announce Type: replace Abstract: Generative policies based on diffusion models and flow matching have shown strong promise for offline reinforcement learning (RL), but their applica

In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores

SafetyDGX agent

arXiv:2605.12530v1 Announce Type: cross Abstract: LLM fairness should be evaluated through in-situ conversational behavior rather than standardized-test Q&A benchmarks. We show that the standardized-t

interwhen: A Generalizable Framework for Steering Reasoning Models with Test-time Verification

SafetyDGX agent

arXiv:2602.11202v3 Announce Type: replace-cross Abstract: Reasoning models produce long traces of intermediate decisions and tool calls, making test-time verification important for ensuring correctnes

Ring-2.6-1T Open sourced today! Soooo looking forward to trying it on Ollama!

Local AiDGX agent

Ring-2.6-1T is a trillion-parameter flagship reasoning model designed for real-world complex task scenarios, now available as an open-source model. The model features about 63B activated parameters pe

13 May 2026

hey surprise - you can just launch interactive in tmux and then tail the jsonl - shipped a small wrapper...ralph loop iterating to full pari…

Model ReleasesDGX agent

hey surprise - you can just launch interactive in tmux and then tail the jsonl - shipped a small wrapper...ralph loop iterating to full parity rn https://github.com/dexhorthy/shannon Starting June 15,

If you use any of the following with your Claude sub, your usage must got cut by 25x: - T3 Code - Conductor - zed - jean - “Claude -p” in yo…

Model ReleasesDGX agent

If you use any of the following with your Claude sub, your usage must got cut by 25x: - T3 Code - Conductor - zed - jean - “Claude -p” in your ci - scripts to call Claude code from other tools They’re

Leveraging RAG for Training-Free Alignment of LLMs

SafetyDGX agent

arXiv:2605.11217v1 Announce Type: new Abstract: Large language model (LLM) alignment algorithms typically consist of post-training over preference pairs. While such algorithms are widely used to enabl

Our continued commitment to Chromebooks, and looking ahead

Model ReleasesDGX agent

At the Android Show yesterday, we introduced Googlebooks: a new category of premium laptops built with Gemini’s helpfulness at the core. Designed for Gemini Intelligence, Googlebooks will give persona

12 May 2026

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2601.01762v2 Announce Type: replace-cross Abstract: Practical autonomous driving requires models that generalize by reasoning through spatial-temporal possibilities to exclude unsafe outcomes. W

AlphaExploitem: Going Beyond the Nash Equilibrium in Poker by Learning to Exploit Suboptimal Play

Model ReleasesDGX agent

arXiv:2605.09150v1 Announce Type: new Abstract: Poker is an imperfect information game that has served as a long-standing benchmark for decision-making under uncertainty. To maximize utility beyond th

Balancing Efficiency and Fairness in Traffic Light Control through Deep Reinforcement Learning

SafetyDGX agent

arXiv:2605.10170v1 Announce Type: new Abstract: Urban traffic congestion presents a significant challenge for modern cities, which impacts mobility and sustainability. Traditional traffic light contro

Communicating Sound Through Natural Language

ResearchDGX agent

arXiv:2605.08750v1 Announce Type: cross Abstract: Natural language is widely used to describe, prompt, and control audio systems, but rarely serves as the representation carrying audio itself. We intr

EFGCL: Learning Dynamic Motion through Spotting-Inspired External Force Guided Curriculum Learning

TutorialsDGX agent

arXiv:2605.10063v1 Announce Type: new Abstract: Learning dynamic whole-body motions for legged robots through reinforcement learning (RL) remains challenging due to the high risk of failure, which mak

Governing AI-Assisted Security Operations: A Design Science Framework for Operational Decision Support

SafetyDGX agent

arXiv:2605.09534v1 Announce Type: cross Abstract: Engineering managers increasingly must decide how to introduce generative artificial intelligence (AI), retrieval-augmented generation, and coding age

Intelligent Autonomous Orchestration for Distributed Cloud Resources using Complex-Stability Analysis

SafetyDGX agent

arXiv:2605.08139v1 Announce Type: cross Abstract: In modern distributed cloud environments, efficient resource allocation is required as traditional scaling mechanisms are often subject to cloud thras

Is Your Driving World Model an All-Around Player?

Model ReleasesDGX agent

arXiv:2605.10858v1 Announce Type: new Abstract: Today's driving world models can generate remarkably realistic dash-cam videos, yet no single model excels universally. Some generate photorealistic tex

LEAF-SQL: Level-wise Exploration with Adaptive Fine-graining for Text-to-SQL Skeleton Prediction

Model ReleasesDGX agent

arXiv:2605.09295v1 Announce Type: new Abstract: Text-to-SQL translates natural language questions into executable SQL queries, enabling intuitive database access for non-experts. While large language

Learning from Trials and Errors: Reflective Test-Time Planning for Embodied LLMs

Model ReleasesDGX agent

arXiv:2602.21198v2 Announce Type: replace-cross Abstract: Embodied LLMs endow robots with high-level task reasoning, but they cannot reflect on what went wrong or why, turning deployment into a sequen

NEXUS: Continual Learning of Symbolic Constraints for Safe and Robust Embodied Planning

SafetyDGX agent

arXiv:2605.09387v1 Announce Type: new Abstract: While Large Language Models (LLMs) have catalyzed progress in embodied intelligence, a fundamental gap between their inherent probabilistic uncertainty

parameter golf was a blast. 2,000+ submissions. 1,000+ verified github accounts. ideas ranging from quantization and depth recurrence to TTT…

Model ReleasesDGX agent

parameter golf was a blast. 2,000+ submissions. 1,000+ verified github accounts. ideas ranging from quantization and depth recurrence to TTT LoRA, SSMs, H-nets, JEPA, and more. autoresearch made itera

ProactBench: Beyond What The User Asked For

Model ReleasesDGX agent

arXiv:2605.09228v1 Announce Type: cross Abstract: Most LLM benchmarks score how well a model responds to explicit requests. They leave unmeasured a different conversational ability: noticing and actin

Reinforcement Learning for Scalable and Trustworthy Intelligent Systems

SafetyDGX agent

arXiv:2605.08378v1 Announce Type: cross Abstract: Reinforcement learning has become a powerful paradigm for improving the capability of intelligent systems, but its practical deployment faces two cent

UniUncer: Unified Dynamic Static Uncertainty for End to End Driving

Local AiDGX agent

arXiv:2603.07686v2 Announce Type: replace-cross Abstract: End-to-end (E2E) driving has become a cornerstone of both industry deployment and academic research, offering a single learnable pipeline that

11 May 2026

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems

ResearchDGX agent

arXiv:2506.04565v2 Announce Type: replace-cross Abstract: Compound AI Systems (CAIS) are an emerging paradigm that integrates large language models (LLMs) with external components, including retriever

GazeVLM: Active Vision via Internal Attention Control for Multimodal Reasoning

Model ReleasesDGX agent

arXiv:2605.07817v1 Announce Type: cross Abstract: Human visual reasoning is governed by active vision, a process where metacognitive control drives top-down goal-directed attention, dynamically routin

Graph Representation Learning Augmented Model Manipulation on Federated Fine-Tuning of LLMs

Model ReleasesDGX agent

arXiv:2605.07961v1 Announce Type: new Abstract: Federated fine-tuning (FFT) has emerged as a privacy-preserving paradigm for collaboratively adapting large language models (LLMs). Built upon federated

← Previous
1…248249250251252…297
Next →