AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,598 results
Safety

Fairness-Aware Multi-Group Target Detection in Online Discussion

DGX agent

arXiv:2407.11933v4 Announce Type: replace Abstract: Target-group detection is the task of detecting which group(s) a piece of content is ``directed at or about''. Applications include targeted marketi

safetyarxiv-cs-lg
23 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

From Fuzzy to Formal: Scaling Hospital Quality Improvement with AI

DGX agent

arXiv:2604.20055v1 Announce Type: new Abstract: Hospital Quality Improvement (QI) plays a critical role in optimizing healthcare delivery by translating high-level hospital goals into actionable solut

safetyarxiv-cs-ai
23 Apr 2026
Safety

Generative Flow Networks for Model Adaptation in Digital Twins of Natural Systems

DGX agent

arXiv:2604.20707v1 Announce Type: new Abstract: Digital twins of natural systems must remain aligned with physical systems that evolve over time, are only partially observed, and are typically modeled

safetyarxiv-cs-lg
23 Apr 2026
Model Releases

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality'…

DGX agent

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality' - Gets lots of stuff done without pausing unnecessarily It

model-releasesdair-ai--x
23 Apr 2026
Safety

Graph2Counsel: Clinically Grounded Synthetic Counseling Dialogue Generation from Client Psychological Graphs

DGX agent

arXiv:2604.20382v1 Announce Type: new Abstract: Rising demand for mental health support has increased interest in using Large Language Models (LLMs) for counseling. However, adapting LLMs to this high

safetyarxiv-cs-cl
23 Apr 2026
Model Releases

HiPO: Hierarchical Preference Optimization for Adaptive Reasoning in LLMs

DGX agent

arXiv:2604.20140v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is an effective framework for aligning large language models with human preferences, but it struggles with complex

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions (Sydney Bradley/Business Insider)

DGX agent

Sydney Bradley / Business Insider: Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions — - In

model-releasestechmeme
23 Apr 2026
Model Releases

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against c…

DGX agent

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against code written using other models Introducing GPT-5.5 A new cla

model-releasessimon-willison--x
23 Apr 2026
Model Releases

KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?

DGX agent

arXiv:2601.13240v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at general programming but struggle with domain-specific software development, necessitating domain special

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m outp…

DGX agent

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m output tokens. - Codex improved its own inference speed 20% lol -

model-releasesswyx--x
23 Apr 2026
Research

Maximum Entropy Semi-Supervised Inverse Reinforcement Learning

DGX agent

arXiv:2604.20074v1 Announce Type: new Abstract: A popular approach to apprenticeship learning (AL) is to formulate it as an inverse reinforcement learning (IRL) problem. The MaxEnt-IRL algorithm succe

researcharxiv-cs-lg
23 Apr 2026
Model Releases

MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror

DGX agent

arXiv:2604.14785v2 Announce Type: replace Abstract: Recent progress in Multimodal Large Language Models (MLLMs) has demonstrated remarkable advances in perception and reasoning, suggesting their poten

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Model Capability Assessment and Safeguards for Biological Weaponization

DGX agent

arXiv:2604.19811v1 Announce Type: cross Abstract: AI leaders and safety reports increasingly warn that advances in model reasoning may enable biological misuse, including by low-expertise users, while

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning

DGX agent

arXiv:2604.20627v1 Announce Type: new Abstract: The temporal lag between actions and their long-term consequences makes credit assignment a challenge when learning goal-directed behaviors from data. G

safetyarxiv-cs-lg
23 Apr 2026
Local Ai

Online Structure Learning and Planning for Autonomous Robot Navigation using Active Inference

DGX agent

arXiv:2510.09574v2 Announce Type: replace Abstract: Autonomous navigation in unfamiliar environments requires robots to simultaneously explore, localise, and plan under uncertainty, without relying on

local-aiarxiv-cs-ro
23 Apr 2026
Model Releases

ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence

DGX agent

arXiv:2604.20719v1 Announce Type: cross Abstract: Omnimodal Notation Processing (ONP) represents a unique frontier for omnimodal AI due to the rigorous, multi-dimensional alignment required across aud

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to fla…

DGX agent

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to flag: there is a certain class of models (one that we’re hitting

model-releasesallie-k--miller--x
23 Apr 2026
Research

OThink-SRR1: Search, Refine and Reasoning with Reinforced Learning for Large Language Models

DGX agent

arXiv:2604.19766v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) expands the knowledge of Large Language Models (LLMs), yet current static retrieval methods struggle with complex

researcharxiv-cs-ai
23 Apr 2026
Model Releases

PR-CAD: Progressive Refinement for Unified Controllable and Faithful Text-to-CAD Generation with Large Language Models

DGX agent

arXiv:2604.19773v1 Announce Type: cross Abstract: The construction of CAD models has traditionally relied on labor-intensive manual operations and specialized expertise. Recent advances in large langu

model-releasesarxiv-cs-ai
23 Apr 2026
Research

R2IF: Aligning Reasoning with Decisions via Composite Rewards for Interpretable LLM Function Calling

DGX agent

arXiv:2604.20316v1 Announce Type: new Abstract: Function calling empowers large language models (LLMs) to interface with external tools, yet existing RL-based approaches suffer from misalignment betwe

researcharxiv-cs-lg
23 Apr 2026
Safety

Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem

DGX agent

arXiv:2604.20805v1 Announce Type: cross Abstract: The value alignment problem for artificial intelligence (AI) is often framed as a purely technical or normative challenge, sometimes focused on hypoth

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

SMARTER: A Data-efficient Framework to Improve Toxicity Detection with Explanation via Self-augmenting Large Language Models

DGX agent

arXiv:2509.15174v3 Announce Type: replace-cross Abstract: WARNING: This paper contains examples of offensive materials. To address the proliferation of toxic content on social media, we introduce SMAR

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

The Ratchet Effect in Silico through Interaction-Driven Cumulative Intelligence in Large Language Models

DGX agent

arXiv:2507.21166v2 Announce Type: replace-cross Abstract: Human intelligence scales through cumulative cultural evolution (CCE), a ratchet process in which innovations are retained against entropic dr

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Things have been degrading super fast in Claude Code. I still use Claude Code, but my default is now Codex. I still prefer Opus models for c…

DGX agent

Things have been degrading super fast in Claude Code. I still use Claude Code, but my default is now Codex. I still prefer Opus models for coding, and so I will try again with the fixes. I appreciate

model-releasesdair-ai--x
23 Apr 2026
Hardware

US Commerce Secretary Howard Lutnick says Nvidia has yet to sell H200 chips to Chinese companies and that the Chinese government has not approved such purchases (Alexandra Alper/Reuters)

DGX agent

Alexandra Alper / Reuters: US Commerce Secretary Howard Lutnick says Nvidia has yet to sell H200 chips to Chinese companies and that the Chinese government has not approved such purchases — Nvidia's (

hardwaretechmeme
23 Apr 2026
Model Releases

WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforcement Learning

DGX agent

arXiv:2604.20398v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel at function-level code generation, project-level tasks such as generating functional and visually aesthetic mul

model-releasesarxiv-cs-cl
23 Apr 2026
Tools

We’re resetting usage limits for subscribers. Thank you so much for your feedback and patience!

DGX agent

Boris Cherny announced that usage limits are being reset for subscribers in response to user feedback and patience. The post suggests this is a policy adjustment made by a service or platform to addre

toolsboris-cherny--x
23 Apr 2026
Tutorials

Accelerating Optimization and Machine Learning through Decentralization

DGX agent

arXiv:2604.19518v1 Announce Type: new Abstract: Decentralized optimization enables multiple devices to learn a global machine learning model while each individual device only has access to its local d

tutorialsarxiv-cs-lg
22 Apr 2026
Model Releases

Announcing Spanner Omni: Your infrastructure, Google’s innovation

DGX agent

Today, we announced the preview of Spanner Omni, a downloadable version of Spanner, that expands its industry-leading distributed database capabilities beyond Google Cloud. This enables enterprises to

model-releasesgoogle-cloud-ai
22 Apr 2026
Safety

ASVSim (AirSim for Surface Vehicles): A High-Fidelity Simulation Framework for Autonomous Surface Vehicle Research

DGX agent

arXiv:2506.22174v2 Announce Type: replace-cross Abstract: The transport industry has recently shown significant interest in unmanned surface vehicles (USVs), specifically for port and inland waterway

safetyarxiv-cs-lg
22 Apr 2026
Research

BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design

DGX agent

arXiv:2508.21184v3 Announce Type: replace-cross Abstract: We propose a general-purpose approach for improving the ability of large language models (LLMs) to intelligently and adaptively gather informa

researcharxiv-cs-ai
22 Apr 2026
Safety

Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications

DGX agent

arXiv:2604.19281v1 Announce Type: cross Abstract: The use of Large Language Models (LLMs) to support patients in addressing medical questions is becoming increasingly prevalent. However, most of the m

safetyarxiv-cs-ai
22 Apr 2026
Safety

Do Emotions Influence Moral Judgment in Large Language Models?

DGX agent

arXiv:2604.19125v1 Announce Type: new Abstract: Large language models have been extensively studied for emotion recognition and moral reasoning as distinct capabilities, yet the extent to which emotio

safetyarxiv-cs-cl
22 Apr 2026
Local Ai

Evaluation-driven Scaling for Scientific Discovery

DGX agent

arXiv:2604.19341v1 Announce Type: cross Abstract: Language models are increasingly used in scientific discovery to generate hypotheses, propose candidate solutions, implement systems, and iteratively

local-aiarxiv-cs-ai
22 Apr 2026
Safety

EVPO: Explained Variance Policy Optimization for Adaptive Critic Utilization in LLM Post-Training

DGX agent

arXiv:2604.19485v1 Announce Type: cross Abstract: Reinforcement learning (RL) for LLM post-training faces a fundamental design choice: whether to use a learned critic as a baseline for policy optimiza

safetyarxiv-cs-ai
22 Apr 2026
Research

GOLD-BEV: GrOund and aeriaL Data for Dense Semantic BEV Mapping of Dynamic Scenes

DGX agent

arXiv:2604.19411v1 Announce Type: cross Abstract: Understanding road scenes in a geometrically consistent, scene-centric representation is crucial for planning and mapping. We present GOLD-BEV, a fram

researcharxiv-cs-ai
22 Apr 2026
Safety

LASER: Learning Active Sensing for Continuum Field Reconstruction

DGX agent

arXiv:2604.19355v1 Announce Type: cross Abstract: High-fidelity measurements of continuum physical fields are essential for scientific discovery and engineering design but remain challenging under spa

safetyarxiv-cs-ai
22 Apr 2026
Safety

Multi-modal Reasoning with LLMs for Visual Semantic Arithmetic

DGX agent

arXiv:2604.19567v1 Announce Type: new Abstract: Reinforcement learning (RL) as post-training is crucial for enhancing the reasoning ability of large language models (LLMs) in coding and math. However,

safetyarxiv-cs-ai
22 Apr 2026
Safety

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge

DGX agent

arXiv:2603.11665v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have been widely adopted as MLLM-as-a-Judges due to their strong alignment with human judgment across vario

safetyarxiv-cs-cl
22 Apr 2026
Safety

Neuromorphic Continual Learning for Sequential Deployment of Nuclear Plant Monitoring Systems

DGX agent

arXiv:2604.18611v1 Announce Type: cross Abstract: Anomaly detection in nuclear industrial control systems (ICS) requires continuous, energy-efficient monitoring across multiple subsystems that are oft

safetyarxiv-cs-ai
22 Apr 2026
Research

OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models

DGX agent

arXiv:2502.16161v2 Announce Type: replace-cross Abstract: Visually-situated text parsing (VsTP) has recently seen notable advancements, driven by the growing demand for automated document understandin

researcharxiv-cs-cl
22 Apr 2026
Model Releases

Paparazzo: Active Mapping of Moving 3D Objects

DGX agent

arXiv:2604.19556v1 Announce Type: new Abstract: Current 3D mapping pipelines generally assume static environments, which limits their ability to accurately capture and reconstruct moving objects. To a

model-releasesarxiv-cs-cv
22 Apr 2026
Research

Reinforcement Learning Enabled Adaptive Multi-Task Control for Bipedal Soccer Robots

DGX agent

arXiv:2604.19104v1 Announce Type: cross Abstract: Developing bipedal football robots in dynamiccombat environments presents challenges related to motionstability and deep coupling of multiple tasks, a

researcharxiv-cs-ai
22 Apr 2026
Safety

RL-ABC: Reinforcement Learning for Accelerator Beamline Control

DGX agent

arXiv:2604.19146v1 Announce Type: new Abstract: Particle accelerator beamline optimization is a high-dimensional control problem traditionally requiring significant expert intervention. We present RLA

safetyarxiv-cs-lg
22 Apr 2026
Model Releases

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation

DGX agent

arXiv:2604.19092v1 Announce Type: cross Abstract: Recent advances in large-scale video world models have enabled increasingly realistic future prediction, raising the prospect of leveraging imagined v

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Semantic Interaction Information mediates compositional generalization in latent space

DGX agent

arXiv:2603.27134v3 Announce Type: replace Abstract: Are there still barriers to generalization once all relevant variables are known? We address this question via a framework that casts compositional

researcharxiv-cs-lg
22 Apr 2026
Industry

SpaceX partners with Cursor on AI training, floats potential $60B acquisition

DGX agent

SpaceX Corp. will help Cursor, a venture-backed vibe coding startup, train artificial intelligence models optimized for programming tasks. The companies announced the initiative on Tuesday. According

industrysiliconangle
22 Apr 2026
Safety

STAR-Teaming: A Strategy-Response Multiplex Network Approach to Automated LLM Red Teaming

DGX agent

arXiv:2604.18976v1 Announce Type: new Abstract: While Large Language Models (LLMs) are widely used, they remain susceptible to jailbreak prompts that can elicit harmful or inappropriate responses. Thi

safetyarxiv-cs-cl
22 Apr 2026
← Previous
1…356357358359360…367
Next →