AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
Model Releases

Transient Turn Injection: Exposing Stateless Multi-Turn Vulnerabilities in Large Language Models

DGX agent

arXiv:2604.21860v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into sensitive workflows, raising the stakes for adversarial robustness and safety. This pape

model-releasesarxiv-cs-ai
24 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Trustworthy Clinical Decision Support Using Meta-Predicates and Domain-Specific Languages

DGX agent

arXiv:2604.21263v1 Announce Type: new Abstract: extbf{Background:} Regulatory frameworks for AI in healthcare, including the EU AI Act and FDA guidance on AI/ML-based medical devices, require clinical

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

When Bigger Isn't Better: A Comprehensive Fairness Evaluation of Political Bias in Multi-News Summarisation

DGX agent

arXiv:2604.21309v1 Announce Type: new Abstract: Multi-document news summarisation systems are increasingly adopted for their convenience in processing vast daily news content, making fairness across d

safetyarxiv-cs-cl
24 Apr 2026
Safety

Why are all LLMs Obsessed with Japanese Culture? On the Hidden Cultural and Regional Biases of LLMs

DGX agent

arXiv:2604.21751v1 Announce Type: cross Abstract: LLMs have been showing limitations when it comes to cultural coverage and competence, and in some cases show regional biases such as amplifying Wester

safetyarxiv-cs-ai
24 Apr 2026
Safety

A Survey of Scaling in Large Language Model Reasoning

DGX agent

arXiv:2504.02181v2 Announce Type: replace Abstract: The rapid advancements in large Language models (LLMs) have significantly enhanced their reasoning capabilities, driven by various strategies such a

safetyarxiv-cs-ai
23 Apr 2026
Research

ALAS: Adaptive Long-Horizon Action Synthesis via Async-pathway Stream Disentanglement

DGX agent

arXiv:2604.20721v1 Announce Type: new Abstract: Long-Horizon (LH) tasks in Human-Scene Interaction (HSI) are complex multi-step tasks that require continuous planning, sequential decision-making, and

researcharxiv-cs-ro
23 Apr 2026
Tutorials

At #RenderCon2026, @EMostaque makes the case during How to Build the Holodeck that stories still matter, IP still matters, but the barriers …

DGX agent

At #RenderCon2026, @EMostaque makes the case during How to Build the Holodeck that stories still matter, IP still matters, but the barriers to building entirely new experiences are dropping faster tha

tutorialsemad-mostaque--x
23 Apr 2026
Model Releases

Automatic Ontology Construction Using LLMs as an External Layer of Memory, Verification, and Planning for Hybrid Intelligent Systems

DGX agent

arXiv:2604.20795v1 Announce Type: new Abstract: This paper presents a hybrid architecture for intelligent systems in which large language models (LLMs) are extended with an external ontological memory

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Breaking the Assistant Mold: Modeling Behavioral Variation in LLM Based Procedural Character Generation

DGX agent

arXiv:2601.03396v2 Announce Type: replace Abstract: Procedural content generation has enabled vast virtual worlds through levels, maps, and quests, but large-scale character generation remains underex

safetyarxiv-cs-cl
23 Apr 2026
Model Releases

CARLA-Air: Fly Drones Inside a CARLA World -- A Unified Infrastructure for Air-Ground Embodied Intelligence

DGX agent

arXiv:2603.28032v2 Announce Type: replace-cross Abstract: The convergence of low-altitude economies, embodied intelligence, and air-ground cooperative systems creates growing demand for simulation inf

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Fairness-Aware Multi-Group Target Detection in Online Discussion

DGX agent

arXiv:2407.11933v4 Announce Type: replace Abstract: Target-group detection is the task of detecting which group(s) a piece of content is ``directed at or about''. Applications include targeted marketi

safetyarxiv-cs-lg
23 Apr 2026
Safety

From Fuzzy to Formal: Scaling Hospital Quality Improvement with AI

DGX agent

arXiv:2604.20055v1 Announce Type: new Abstract: Hospital Quality Improvement (QI) plays a critical role in optimizing healthcare delivery by translating high-level hospital goals into actionable solut

safetyarxiv-cs-ai
23 Apr 2026
Safety

Generative Flow Networks for Model Adaptation in Digital Twins of Natural Systems

DGX agent

arXiv:2604.20707v1 Announce Type: new Abstract: Digital twins of natural systems must remain aligned with physical systems that evolve over time, are only partially observed, and are typically modeled

safetyarxiv-cs-lg
23 Apr 2026
Model Releases

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality'…

DGX agent

GPT-5.5 in Codex is a delight to work with: - Super sharp with responses - It understands intent better than any model - Great 'personality' - Gets lots of stuff done without pausing unnecessarily It

model-releasesdair-ai--x
23 Apr 2026
Safety

Graph2Counsel: Clinically Grounded Synthetic Counseling Dialogue Generation from Client Psychological Graphs

DGX agent

arXiv:2604.20382v1 Announce Type: new Abstract: Rising demand for mental health support has increased interest in using Large Language Models (LLMs) for counseling. However, adapting LLMs to this high

safetyarxiv-cs-cl
23 Apr 2026
Model Releases

HiPO: Hierarchical Preference Optimization for Adaptive Reasoning in LLMs

DGX agent

arXiv:2604.20140v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is an effective framework for aligning large language models with human preferences, but it struggles with complex

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions (Sydney Bradley/Business Insider)

DGX agent

Sydney Bradley / Business Insider: Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions — - In

model-releasestechmeme
23 Apr 2026
Model Releases

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against c…

DGX agent

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against code written using other models Introducing GPT-5.5 A new cla

model-releasessimon-willison--x
23 Apr 2026
Model Releases

KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?

DGX agent

arXiv:2601.13240v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at general programming but struggle with domain-specific software development, necessitating domain special

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m outp…

DGX agent

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m output tokens. - Codex improved its own inference speed 20% lol -

model-releasesswyx--x
23 Apr 2026
Research

Maximum Entropy Semi-Supervised Inverse Reinforcement Learning

DGX agent

arXiv:2604.20074v1 Announce Type: new Abstract: A popular approach to apprenticeship learning (AL) is to formulate it as an inverse reinforcement learning (IRL) problem. The MaxEnt-IRL algorithm succe

researcharxiv-cs-lg
23 Apr 2026
Model Releases

MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror

DGX agent

arXiv:2604.14785v2 Announce Type: replace Abstract: Recent progress in Multimodal Large Language Models (MLLMs) has demonstrated remarkable advances in perception and reasoning, suggesting their poten

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Model Capability Assessment and Safeguards for Biological Weaponization

DGX agent

arXiv:2604.19811v1 Announce Type: cross Abstract: AI leaders and safety reports increasingly warn that advances in model reasoning may enable biological misuse, including by low-expertise users, while

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning

DGX agent

arXiv:2604.20627v1 Announce Type: new Abstract: The temporal lag between actions and their long-term consequences makes credit assignment a challenge when learning goal-directed behaviors from data. G

safetyarxiv-cs-lg
23 Apr 2026
Local Ai

Online Structure Learning and Planning for Autonomous Robot Navigation using Active Inference

DGX agent

arXiv:2510.09574v2 Announce Type: replace Abstract: Autonomous navigation in unfamiliar environments requires robots to simultaneously explore, localise, and plan under uncertainty, without relying on

local-aiarxiv-cs-ro
23 Apr 2026
Model Releases

ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence

DGX agent

arXiv:2604.20719v1 Announce Type: cross Abstract: Omnimodal Notation Processing (ONP) represents a unique frontier for omnimodal AI due to the rigorous, multi-dimensional alignment required across aud

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to fla…

DGX agent

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to flag: there is a certain class of models (one that we’re hitting

model-releasesallie-k--miller--x
23 Apr 2026
Research

OThink-SRR1: Search, Refine and Reasoning with Reinforced Learning for Large Language Models

DGX agent

arXiv:2604.19766v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) expands the knowledge of Large Language Models (LLMs), yet current static retrieval methods struggle with complex

researcharxiv-cs-ai
23 Apr 2026
Model Releases

PR-CAD: Progressive Refinement for Unified Controllable and Faithful Text-to-CAD Generation with Large Language Models

DGX agent

arXiv:2604.19773v1 Announce Type: cross Abstract: The construction of CAD models has traditionally relied on labor-intensive manual operations and specialized expertise. Recent advances in large langu

model-releasesarxiv-cs-ai
23 Apr 2026
Research

R2IF: Aligning Reasoning with Decisions via Composite Rewards for Interpretable LLM Function Calling

DGX agent

arXiv:2604.20316v1 Announce Type: new Abstract: Function calling empowers large language models (LLMs) to interface with external tools, yet existing RL-based approaches suffer from misalignment betwe

researcharxiv-cs-lg
23 Apr 2026
Safety

Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem

DGX agent

arXiv:2604.20805v1 Announce Type: cross Abstract: The value alignment problem for artificial intelligence (AI) is often framed as a purely technical or normative challenge, sometimes focused on hypoth

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

SMARTER: A Data-efficient Framework to Improve Toxicity Detection with Explanation via Self-augmenting Large Language Models

DGX agent

arXiv:2509.15174v3 Announce Type: replace-cross Abstract: WARNING: This paper contains examples of offensive materials. To address the proliferation of toxic content on social media, we introduce SMAR

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

The Ratchet Effect in Silico through Interaction-Driven Cumulative Intelligence in Large Language Models

DGX agent

arXiv:2507.21166v2 Announce Type: replace-cross Abstract: Human intelligence scales through cumulative cultural evolution (CCE), a ratchet process in which innovations are retained against entropic dr

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Things have been degrading super fast in Claude Code. I still use Claude Code, but my default is now Codex. I still prefer Opus models for c…

DGX agent

Things have been degrading super fast in Claude Code. I still use Claude Code, but my default is now Codex. I still prefer Opus models for coding, and so I will try again with the fixes. I appreciate

model-releasesdair-ai--x
23 Apr 2026
Hardware

US Commerce Secretary Howard Lutnick says Nvidia has yet to sell H200 chips to Chinese companies and that the Chinese government has not approved such purchases (Alexandra Alper/Reuters)

DGX agent

Alexandra Alper / Reuters: US Commerce Secretary Howard Lutnick says Nvidia has yet to sell H200 chips to Chinese companies and that the Chinese government has not approved such purchases — Nvidia's (

hardwaretechmeme
23 Apr 2026
Model Releases

WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforcement Learning

DGX agent

arXiv:2604.20398v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel at function-level code generation, project-level tasks such as generating functional and visually aesthetic mul

model-releasesarxiv-cs-cl
23 Apr 2026
Tools

We’re resetting usage limits for subscribers. Thank you so much for your feedback and patience!

DGX agent

Boris Cherny announced that usage limits are being reset for subscribers in response to user feedback and patience. The post suggests this is a policy adjustment made by a service or platform to addre

toolsboris-cherny--x
23 Apr 2026
Tutorials

Accelerating Optimization and Machine Learning through Decentralization

DGX agent

arXiv:2604.19518v1 Announce Type: new Abstract: Decentralized optimization enables multiple devices to learn a global machine learning model while each individual device only has access to its local d

tutorialsarxiv-cs-lg
22 Apr 2026
Model Releases

Announcing Spanner Omni: Your infrastructure, Google’s innovation

DGX agent

Today, we announced the preview of Spanner Omni, a downloadable version of Spanner, that expands its industry-leading distributed database capabilities beyond Google Cloud. This enables enterprises to

model-releasesgoogle-cloud-ai
22 Apr 2026
Safety

ASVSim (AirSim for Surface Vehicles): A High-Fidelity Simulation Framework for Autonomous Surface Vehicle Research

DGX agent

arXiv:2506.22174v2 Announce Type: replace-cross Abstract: The transport industry has recently shown significant interest in unmanned surface vehicles (USVs), specifically for port and inland waterway

safetyarxiv-cs-lg
22 Apr 2026
Research

BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design

DGX agent

arXiv:2508.21184v3 Announce Type: replace-cross Abstract: We propose a general-purpose approach for improving the ability of large language models (LLMs) to intelligently and adaptively gather informa

researcharxiv-cs-ai
22 Apr 2026
Safety

Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications

DGX agent

arXiv:2604.19281v1 Announce Type: cross Abstract: The use of Large Language Models (LLMs) to support patients in addressing medical questions is becoming increasingly prevalent. However, most of the m

safetyarxiv-cs-ai
22 Apr 2026
Safety

Do Emotions Influence Moral Judgment in Large Language Models?

DGX agent

arXiv:2604.19125v1 Announce Type: new Abstract: Large language models have been extensively studied for emotion recognition and moral reasoning as distinct capabilities, yet the extent to which emotio

safetyarxiv-cs-cl
22 Apr 2026
Local Ai

Evaluation-driven Scaling for Scientific Discovery

DGX agent

arXiv:2604.19341v1 Announce Type: cross Abstract: Language models are increasingly used in scientific discovery to generate hypotheses, propose candidate solutions, implement systems, and iteratively

local-aiarxiv-cs-ai
22 Apr 2026
Safety

EVPO: Explained Variance Policy Optimization for Adaptive Critic Utilization in LLM Post-Training

DGX agent

arXiv:2604.19485v1 Announce Type: cross Abstract: Reinforcement learning (RL) for LLM post-training faces a fundamental design choice: whether to use a learned critic as a baseline for policy optimiza

safetyarxiv-cs-ai
22 Apr 2026
Research

GOLD-BEV: GrOund and aeriaL Data for Dense Semantic BEV Mapping of Dynamic Scenes

DGX agent

arXiv:2604.19411v1 Announce Type: cross Abstract: Understanding road scenes in a geometrically consistent, scene-centric representation is crucial for planning and mapping. We present GOLD-BEV, a fram

researcharxiv-cs-ai
22 Apr 2026
Safety

LASER: Learning Active Sensing for Continuum Field Reconstruction

DGX agent

arXiv:2604.19355v1 Announce Type: cross Abstract: High-fidelity measurements of continuum physical fields are essential for scientific discovery and engineering design but remain challenging under spa

safetyarxiv-cs-ai
22 Apr 2026
Safety

Multi-modal Reasoning with LLMs for Visual Semantic Arithmetic

DGX agent

arXiv:2604.19567v1 Announce Type: new Abstract: Reinforcement learning (RL) as post-training is crucial for enhancing the reasoning ability of large language models (LLMs) in coding and math. However,

safetyarxiv-cs-ai
22 Apr 2026
← Previous
1…356357358359360…367
Next →