AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,360 results
Safety

CARL: Criticality-Aware Agentic Reinforcement Learning

DGX agent

arXiv:2512.04949v3 Announce Type: replace-cross Abstract: Agents capable of accomplishing complex tasks through multiple interactions with the environment have emerged as a popular research direction.

safetyarxiv-cs-ai
12 May 2026
Safety

Compute Where it Counts: Self Optimizing Language Models

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.10875v1 Announce Type: cross Abstract: Efficient LLM inference research has largely focused on reducing the cost of each decoding step (e.g., using quantization, pruning, or sparse attentio

safetyarxiv-cs-cl
12 May 2026
Safety

Conformity Generates Collective Misalignment in AI Agents Societies

DGX agent

arXiv:2605.10721v1 Announce Type: cross Abstract: Artificial intelligence safety research focuses on aligning individual language models with human values, yet deployed AI systems increasingly operate

safetyarxiv-cs-cl
12 May 2026
Safety

Embodied AI in Action: Insights from SAE World Congress 2026 on Safety, Trust, Robotics, and Real-World Deployment

DGX agent

arXiv:2605.10653v1 Announce Type: new Abstract: Embodied artificial intelligence is rapidly moving from research into real-world systems such as autonomous vehicles, mobile robots, and industrial mach

safetyarxiv-cs-ro
12 May 2026
Model Releases

Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition

DGX agent

arXiv:2605.10127v1 Announce Type: new Abstract: Recent research work on fashion outfit generation focuses on promoting visual consistency of garments by leveraging key information from reference image

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

LegalCiteBench: Evaluating Citation Reliability in Legal Language Models

DGX agent

arXiv:2605.10186v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into legal drafting and research workflows, where incorrect citations or fabricated precedent

model-releasesarxiv-cs-ai
12 May 2026
Applications

LLARS: Enabling Domain Expert & Developer Collaboration for LLM Prompting, Generation and Evaluation

DGX agent

arXiv:2605.10593v1 Announce Type: new Abstract: We demonstrate LLARS (LLM Assisted Research System), an open-source platform that bridges the gap between domain experts and developers for building LLM

applicationsarxiv-cs-ai
12 May 2026
Applications

Marrying Generative Model of Healthcare Events with Digital Twin of Social Determinants of Health for Disease Reasoning

DGX agent

arXiv:2605.09771v1 Announce Type: new Abstract: Despite the central role of sensor-derived measurements such as imaging traits and plasma biomarkers in biomedical research and clinical practice, exist

applicationsarxiv-cs-ai
12 May 2026
Model Releases

MOTOR-Bench: A Real-world Dataset and Multi-agent Framework for Zero-shot Human Mental State Understanding

DGX agent

arXiv:2605.09703v1 Announce Type: new Abstract: Understanding human mental states from natural behavior is crucial for intelligent systems in the real world. However, most current research focuses on

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Navigating the Sea of LLM Evaluation: Investigating Bias in Toxicity Benchmarks

DGX agent

arXiv:2605.10639v1 Announce Type: new Abstract: The rapid adoption of LLMs in both research and industry highlights the challenges of deploying them safely and reveals a gap in the systematic evaluati

model-releasesarxiv-cs-ai
12 May 2026
Agents

NyayaAI: An AI-Powered Legal Assistant Using Multi-Agent Architecture and Retrieval-Augmented Generation

DGX agent

arXiv:2605.10155v1 Announce Type: new Abstract: Legal information in India remains largely inaccessible due to the complexity of legal language and the sheer volume of legal documentation involved in

agentsarxiv-cs-cl
12 May 2026
Model Releases

Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization

DGX agent

arXiv:2605.09996v1 Announce Type: new Abstract: While multimodal large language models have advanced across text, image, and audio, personalization research has remained primarily vision-language, wit

model-releasesarxiv-cs-cv
12 May 2026
Safety

PersonaTeaming: Supporting Persona-Driven Red-Teaming for Generative AI

DGX agent

arXiv:2605.05682v2 Announce Type: replace-cross Abstract: Recent developments in AI safety research have called for red-teaming methods that effectively surface potential risks posed by generative AI

safetyarxiv-cs-ai
12 May 2026
Safety

Political Plasticity: An Analysis of Ideological Adaptability in Large Language Models

DGX agent

arXiv:2605.08415v1 Announce Type: new Abstract: Since the advent of Large Language Models (LLMs), a significant area of research has focused on their intrinsic biases, particularly in political discou

safetyarxiv-cs-ai
12 May 2026
Model Releases

PumpSense: Real-Time Detection and Target Extraction of Crypto Pump-and-Dumps on Telegram

DGX agent

arXiv:2605.09431v1 Announce Type: new Abstract: Cryptocurrency pump-and-dump schemes coordinated via Telegram threaten market integrity. However, existing research addressing this specific threat has

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Rethinking Agentic Search with Pi-Serini: Is Lexical Retrieval Sufficient?

DGX agent

arXiv:2605.10848v1 Announce Type: cross Abstract: Does a lexical retriever suffice as large language models (LLMs) become more capable in an agentic loop? This question naturally arises when building

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SciIntegrity-Bench: A Benchmark for Evaluating Academic Integrity in AI Scientist Systems

DGX agent

arXiv:2605.10246v1 Announce Type: new Abstract: AI scientist systems are increasingly deployed for autonomous research, yet their academic integrity has never been systematically evaluated. We introdu

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Single-Configuration Attack Success Rate Is Not Enough: Jailbreak Evaluations Should Report Distributional Attack Success

DGX agent

arXiv:2605.09070v1 Announce Type: cross Abstract: Many jailbreak attack research papers report attack success rates for a limited number of parameter settings, even though there are many combinations

model-releasesarxiv-cs-ai
12 May 2026
Agents

SoK: A Systematic Bidirectional Literature Review of AI & DLT Convergence

DGX agent

arXiv:2605.10515v1 Announce Type: cross Abstract: The integration of Artificial Intelligence (AI) with Distributed Ledger Technology (DLT) has become a growing research area, yet contributions tend to

agentsarxiv-cs-ai
12 May 2026
Model Releases

Teaching LLMs to See Graphs: Unifying Text and Structural Reasoning

DGX agent

arXiv:2605.10247v1 Announce Type: new Abstract: Using Large Language Models (LLMs) to process graph-structured data is an active research area, yet current state-of-the-art approaches typically rely o

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

UFO: A Unified Flow-Oriented Framework for Robust Continual Graph Learning

DGX agent

arXiv:2605.09862v1 Announce Type: cross Abstract: Graph learning research has increasingly shifted toward continual graph learning (CGL), which better reflects real-world scenarios where graphs evolve

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

UniUncer: Unified Dynamic Static Uncertainty for End to End Driving

DGX agent

arXiv:2603.07686v2 Announce Type: replace-cross Abstract: End-to-end (E2E) driving has become a cornerstone of both industry deployment and academic research, offering a single learnable pipeline that

local-aiarxiv-cs-cv
12 May 2026
Agents

What Software Engineering Looks Like to AI Agents? -- An Empirical Study of AI-Only Technical Discourse on MoltBook

DGX agent

arXiv:2605.08380v1 Announce Type: cross Abstract: AI agents are increasingly framed as software-engineering teammates, yet most research studies them inside human-centered workflows. Little is known a

agentsarxiv-cs-ai
12 May 2026
Agents

AgentProg: Empowering Long-Horizon GUI Agents with Program-Guided Context Management

DGX agent

arXiv:2512.10371v2 Announce Type: replace Abstract: The rapid development of mobile GUI agents has stimulated growing research interest in long-horizon task automation. However, building agents for th

agentsarxiv-cs-ai
11 May 2026
Model Releases

AI CFD Scientist: Toward Open-Ended Computational Fluid Dynamics Discovery with Physics-Aware AI Agents

DGX agent

arXiv:2605.06607v2 Announce Type: replace-cross Abstract: Recent LLM-based agents have closed substantial portions of the scientific discovery loop in software-only machine-learning research, in chemi

model-releasesarxiv-cs-ai
11 May 2026
Agents

Are LLM Agents Behaviorally Coherent? Latent Profiles for Social Simulation

DGX agent

arXiv:2509.03736v2 Announce Type: replace Abstract: The impressive capabilities of Large Language Models (LLMs) raise the possibility that synthetic agents can serve as substitutes for real participan

agentsarxiv-cs-ai
11 May 2026
Safety

BEAVER: An Efficient Deterministic LLM Verifier

DGX agent

arXiv:2512.05439v2 Announce Type: replace Abstract: As large language models (LLMs) transition from research prototypes to production systems, practitioners often need reliable methods to verify model

safetyarxiv-cs-ai
11 May 2026
Model Releases

FAME: Forecasting Academic Impact via Continuous-Time Manifold Evolution

DGX agent

arXiv:2605.07208v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to brainstorm and evaluate research ideas, yet assessing such judgments is fundamentally difficult be

model-releasesarxiv-cs-lg
11 May 2026
Safety

Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment

DGX agent

arXiv:2605.08064v1 Announce Type: new Abstract: Spatial intelligence in vision-language models (VLMs) attracts research interest with the practical demand to reason in the 3D world.Despite promising r

safetyarxiv-cs-cv
11 May 2026
Safety

Robustness of Refugee-Matching Gains to Off-Policy Evaluation Choices

DGX agent

arXiv:2605.06686v1 Announce Type: new Abstract: Previous research has investigated the potential of refugee matching for boosting refugee outcomes, first considered by Bansak et al. (2018). This paper

safetyarxiv-cs-lg
11 May 2026
Safety

The Proxy Presumption: From Semantic Embeddings to Valid Social Measures

DGX agent

arXiv:2605.07409v1 Announce Type: new Abstract: Natural Language Processing is rapidly evolving into a primary instrument for Computational Social Science, with researchers increasingly using embeddin

safetyarxiv-cs-cl
11 May 2026
Safety

Think-with-Rubrics: From External Evaluator to Internal Reasoning Guidance

DGX agent

arXiv:2605.07461v1 Announce Type: new Abstract: Rubrics have been extensively utilized for evaluating unverifiable, open-ended tasks, with recent research incorporating them into reward systems for re

safetyarxiv-cs-cl
11 May 2026
Agents

VDCook:DIY video data cook your MLLMs

DGX agent

arXiv:2603.05539v2 Announce Type: replace-cross Abstract: We introduce VDCook: a self-evolving video data operating system, a configurable video data construction platform for researchers and vertical

agentsarxiv-cs-ai
11 May 2026
Agents

WebClipper: Efficient Evolution of Web Agents with Graph-based Trajectory Pruning

DGX agent

arXiv:2602.12852v2 Announce Type: replace Abstract: Deep Research systems based on web agents have shown strong potential in solving complex information-seeking tasks, yet their search efficiency rema

agentsarxiv-cs-ai
11 May 2026
Agents

AI-Aided Advancements in Autonomous Underwater Vehicle Navigation

DGX agent

arXiv:2605.04672v1 Announce Type: new Abstract: Autonomous underwater vehicles (AUVs) have become indispensable for deep-sea exploration, spanning critical scientific research and commercial applicati

agentsarxiv-cs-ro
7 May 2026
Tutorials

Forget BIT, It is All about TOKEN: Towards Semantic Information Theory for LLMs

DGX agent

arXiv:2511.01202v3 Announce Type: replace-cross Abstract: Despite the unprecedented empirical triumphs of LLMs across diverse real-world applications, the prevailing research paradigm remains overwhel

tutorialsarxiv-cs-ai
7 May 2026
Safety

Graph-Augmented LLMs for Swiss MP Ideology Prediction

DGX agent

arXiv:2605.04643v1 Announce Type: new Abstract: Approximating the ideological position of Members of Parliament (MPs) is a fundamental task in political science, helping researchers understand legisla

safetyarxiv-cs-cl
7 May 2026
Applications

When LLMs get significantly worse: A statistical approach to detect model degradations

DGX agent

arXiv:2602.10144v2 Announce Type: replace-cross Abstract: Minimizing the inference cost and latency of foundation models has become a crucial area of research. Optimization approaches include theoreti

applicationsarxiv-cs-lg
7 May 2026
Tutorials

Cripping AI: Reimagining AI Through Lived Disability Experiences

DGX agent

arXiv:2605.02080v1 Announce Type: cross Abstract: Drawing on crip theory, this paper proposes cripping AI as a guiding framework to center lived disability experiences in AI research and development.

tutorialsarxiv-cs-ai
6 May 2026
Tutorials

Enhance the after-discharge mortality rate prediction via learning from the medical notes

DGX agent

arXiv:2605.03560v1 Announce Type: new Abstract: With the increase of the Electronic Health Records (EHR) data, more and more researchers are developing machine learning models to learn from the medica

tutorialsarxiv-cs-lg
6 May 2026
Hardware

MOOSE-Star: Unlocking Tractable Training for Scientific Discovery by Breaking the Complexity Barrier

DGX agent

arXiv:2603.03756v3 Announce Type: replace-cross Abstract: While large language models (LLMs) show promise in scientific discovery, existing research focuses on inference or feedback-driven training, l

hardwarearxiv-cs-cl
6 May 2026
Model Releases

Reward Hacking Benchmark: Measuring Exploits in LLM Agents with Tool Use

DGX agent

arXiv:2605.02964v1 Announce Type: new Abstract: Reinforcement learning (RL) trained language model agents with tool access are increasingly deployed in coding assistants, research tools, and autonomou

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Seeking Information with RAG-Assistants: Does Model Size Matter in Human-AI Collaborations?

DGX agent

arXiv:2605.00964v1 Announce Type: cross Abstract: Much research on LLMs has focused on increasing benchmark performance. However, the evaluation of such models in real-world collaborative human-AI wor

model-releasesarxiv-cs-ai
6 May 2026
Safety

SoDa2: Single-Stage Open-Set Domain Adaptation via Decoupled Alignment for Cross-Scene Hyperspectral Image Classification

DGX agent

arXiv:2605.03371v1 Announce Type: new Abstract: Cross-scene hyperspectral image (HSI) classification stands as a fundamental research topic in remote sensing, with extensive applications spanning vari

safetyarxiv-cs-cv
6 May 2026
Model Releases

Towards Agentic Runtime Healing

DGX agent

arXiv:2408.01055v2 Announce Type: replace-cross Abstract: Self-healing systems have long been a focus of research, aiming to enable software to recover from unexpected runtime errors without human int

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Towards Understanding Specification Gaming in Reasoning Models

DGX agent

arXiv:2605.02269v1 Announce Type: new Abstract: Specification gaming is a critical failure mode of LLM agents. Despite this, there has been little systematic research into when it arises and what driv

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Aligning LLMs with Biomedical Knowledge using Balanced Fine-Tuning

DGX agent

arXiv:2511.21075v3 Announce Type: replace Abstract: Engineering LLMs to accelerate life sciences research requires a robust alignment with biomedical knowledge. We observe that biomedical text exhibit

model-releasesarxiv-cs-lg
5 May 2026
Safety

Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure

DGX agent

arXiv:2605.00055v1 Announce Type: cross Abstract: We report a safety incident in a deployed multi-agent research system in which a primary AI agent installed 107 unauthorized software components, over

safetyarxiv-cs-ai
5 May 2026
← Previous
1…396397398399400…466
Next →