AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
21,937 results
Model Releases

SEW: Self-Evolving Agentic Workflows for Automated Code Generation

DGX agent

arXiv:2505.18646v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated effectiveness in code generation tasks. To enable LLMs to address more complex coding challenge

model-releasesarxiv-cs-ai
15 Apr 2026
Tutorials
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Siamese Foundation Models for Crystal Structure Prediction

DGX agent

arXiv:2503.10471v2 Announce Type: replace-cross Abstract: Predicting crystal structures from chemical compositions is a fundamental challenge in materials discovery, complicated by complex 3D geometri

tutorialsarxiv-cs-ai
15 Apr 2026
Model Releases

SIRI-Bench: Challenging VLMs' Spatial Intelligence through Complex Reasoning Tasks

DGX agent

arXiv:2506.14512v4 Announce Type: replace Abstract: Large Language Models (LLMs) have undergone rapid progress, largely attributed to reinforcement learning on complex reasoning tasks. In contrast, wh

model-releasesarxiv-cs-cv
15 Apr 2026
Agents

Thermodynamic Liquid Manifold Networks: Physics-Bounded Deep Learning for Solar Forecasting in Autonomous Off-Grid Microgrids

DGX agent

arXiv:2604.11909v1 Announce Type: cross Abstract: The stable operation of autonomous off-grid photovoltaic systems requires solar forecasting algorithms that respect atmospheric thermodynamics. Contem

agentsarxiv-cs-ai
15 Apr 2026
Model Releases

TRUST Agents: A Collaborative Multi-Agent Framework for Fake News Detection, Explainable Verification, and Logic-Aware Claim Reasoning

DGX agent

arXiv:2604.12184v1 Announce Type: new Abstract: TRUST Agents is a collaborative multi-agent framework for explainable fact verification and fake news detection. Rather than treating verification as a

model-releasesarxiv-cs-ai
15 Apr 2026
Applications

Uncertainty Quantification on Graph Learning: A Survey

DGX agent

arXiv:2404.14642v4 Announce Type: replace Abstract: Graphical models have demonstrated their exceptional capabilities across numerous applications. However, their performance, confidence, and trustwor

applicationsarxiv-cs-lg
15 Apr 2026
Safety

WebChain: A Large-Scale Human-Annotated Dataset of Real-World Web Interaction Traces

DGX agent

arXiv:2603.05295v3 Announce Type: replace Abstract: We introduce WebChain, the largest open-source dataset of human-annotated trajectories on real-world websites, designed to accelerate reproducible r

safetyarxiv-cs-ai
15 Apr 2026
Tutorials

Why Did Apple Fall: Evaluating Curiosity in Large Language Models

DGX agent

arXiv:2510.20635v2 Announce Type: replace-cross Abstract: Curiosity serves as a pivotal conduit for human beings to discover and learn new knowledge. Recent advancements of large language models (LLMs

tutorialsarxiv-cs-ai
15 Apr 2026
Model Releases

ACE-Bench: A Lightweight Benchmark for Evaluating Azure SDK Usage Correctness

DGX agent

arXiv:2604.09564v1 Announce Type: cross Abstract: We present ACE-Bench (Azure SDK Coding Evaluation Benchmark), an execution-free benchmark that provides fast, reproducible pass or fail signals for wh

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Agentic Exploration of PDE Spaces using Latent Foundation Models for Parameterized Simulations

DGX agent

arXiv:2604.09584v1 Announce Type: new Abstract: Flow physics and more broadly physical phenomena governed by partial differential equations (PDEs), are inherently continuous, high-dimensional and ofte

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

AI Integrity: A New Paradigm for Verifiable AI Governance

DGX agent

arXiv:2604.11065v1 Announce Type: new Abstract: AI systems increasingly shape high-stakes decisions in healthcare, law, defense, and education, yet existing governance paradigms -- AI Ethics, AI Safet

safetyarxiv-cs-ai
14 Apr 2026
Safety

Assessing Model-Agnostic XAI Methods against EU AI Act Explainability Requirements

DGX agent

arXiv:2604.09628v1 Announce Type: cross Abstract: Explainable AI (XAI) has evolved in response to expectations and regulations, such as the EU AI Act, which introduces regulatory requirements on AI-po

safetyarxiv-cs-ai
14 Apr 2026
Applications

Automatic Uncertainty-Aware Synthetic Data Bootstrapping for Historical Map Segmentation

DGX agent

arXiv:2511.15875v2 Announce Type: replace Abstract: The automated analysis of historical documents, particularly maps, has drastically benefited from advances in deep learning and its success across v

applicationsarxiv-cs-cv
14 Apr 2026
Applications

BEM: Training-Free Background Embedding Memory for False-Positive Suppression in Real-Time Fixed-Background Camera

DGX agent

arXiv:2604.11714v1 Announce Type: new Abstract: Pretrained detectors perform well on benchmarks but often suffer performance degradation in real-world deployments due to distribution gaps between trai

applicationsarxiv-cs-cv
14 Apr 2026
Applications

Beyond Fixed False Discovery Rates: Post-Hoc Conformal Selection with E-Variables

DGX agent

arXiv:2604.11305v1 Announce Type: new Abstract: Conformal selection (CS) uses calibration data to identify test inputs whose unobserved outcomes are likely to satisfy a pre-specified minimal quality r

applicationsarxiv-cs-lg
14 Apr 2026
Safety

Beyond Message Passing: A Semantic View of Agent Communication Protocols

DGX agent

arXiv:2604.02369v3 Announce Type: replace-cross Abstract: Agent communication protocols are becoming critical infrastructure for large language model (LLM) systems that must use tools, coordinate with

safetyarxiv-cs-ai
14 Apr 2026
Applications

Bootstrapping Video Semantic Segmentation Model via Distillation-assisted Test-Time Adaptation

DGX agent

arXiv:2604.10950v1 Announce Type: new Abstract: Fully supervised Video Semantic Segmentation (VSS) relies heavily on densely annotated video data, limiting practical applicability. Alternatively, appl

applicationsarxiv-cs-cv
14 Apr 2026
Model Releases

C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts

DGX agent

arXiv:2604.11796v1 Announce Type: cross Abstract: Recently, large language models (LLMs) are capable of generating highly fluent textual content. While they offer significant convenience to humans, th

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Can Large Language Models Infer Causal Relationships from Real-World Text?

DGX agent

arXiv:2505.18931v4 Announce Type: replace Abstract: Understanding and inferring causal relationships from texts is a core aspect of human cognition and is essential for advancing large language models

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

CapyMOA: Efficient Machine Learning for Data Streams and Online Continual Learning in Python

DGX agent

arXiv:2502.07432v2 Announce Type: replace Abstract: CapyMOA is an open-source Python library for efficient machine learning on data streams and online continual learning. It provides a structured fram

tutorialsarxiv-cs-lg
14 Apr 2026
Model Releases

Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models

DGX agent

arXiv:2503.21380v3 Announce Type: replace Abstract: The rapid advancement of large reasoning models has saturated existing math benchmarks, underscoring the urgent need for more challenging evaluation

model-releasesarxiv-cs-cl
14 Apr 2026
Agents

ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents

DGX agent

arXiv:2509.22830v3 Announce Type: replace Abstract: The growing deployment of large language model (LLM) based agents that interact with external environments has created new attack surfaces for adver

agentsarxiv-cs-cl
14 Apr 2026
Model Releases

Comparative Analysis of Large Language Models in Healthcare

DGX agent

arXiv:2604.10316v1 Announce Type: new Abstract: Background: Large Language Models (LLMs) are transforming artificial intelligence applications in healthcare due to their ability to understand, generat

model-releasesarxiv-cs-cl
14 Apr 2026
Agents

Deep Learning for Sequential Decision Making under Uncertainty: Foundations, Frameworks, and Frontiers

DGX agent

arXiv:2604.11507v1 Announce Type: cross Abstract: Artificial intelligence (AI) is moving increasingly beyond prediction to support decisions in complex, uncertain, and dynamic environments. This shift

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

DiningBench: A Hierarchical Multi-view Benchmark for Perception and Reasoning in the Dietary Domain

DGX agent

arXiv:2604.10425v1 Announce Type: new Abstract: Recent advancements in Vision-Language Models (VLMs) have revolutionized general visual understanding. However, their application in the food domain rem

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Disambiguation-Centric Finetuning Makes Enterprise Tool-Calling LLMs More Realistic and Less Risky

DGX agent

arXiv:2507.03336v4 Announce Type: replace Abstract: Large language models (LLMs) are increasingly tasked with invoking enterprise APIs, yet they routinely falter when near-duplicate tools vie for the

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Distributionally Robust PAC-Bayesian Control

DGX agent

arXiv:2604.10588v1 Announce Type: new Abstract: We present a distributionally robust PAC-Bayesian framework for certifying the performance of learning-based finite-horizon controllers. While existing

safetyarxiv-cs-lg
14 Apr 2026
Model Releases

Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models

DGX agent

arXiv:2601.03926v2 Announce Type: replace Abstract: The deployment of Large Vision-Language Models (LVLMs) for real-world document question answering is often constrained by dynamic, user-defined poli

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

DocRevive: A Unified Pipeline for Document Text Restoration

DGX agent

arXiv:2604.10077v1 Announce Type: new Abstract: In Document Understanding, the challenge of reconstructing damaged, occluded, or incomplete text remains a critical yet unexplored problem. Subsequent d

model-releasesarxiv-cs-cv
14 Apr 2026
Applications

Does Your VFM Speak Plant? The Botanical Grammar of Vision Foundation Models for Object Detection

DGX agent

arXiv:2604.09920v1 Announce Type: new Abstract: Vision foundation models (VFMs) offer the promise of zero-shot object detection without task-specific training data, yet their performance in complex ag

applicationsarxiv-cs-cv
14 Apr 2026
Applications

Domain-Specific Data Generation Framework for RAG Adaptation

DGX agent

arXiv:2510.11217v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) combines the language understanding and reasoning power of large language models (LLMs) with external ret

applicationsarxiv-cs-ai
14 Apr 2026
Applications

DoSReMC: Domain Shift Resilient Mammography Classification using Batch Normalization Adaptation

DGX agent

arXiv:2508.15452v3 Announce Type: replace-cross Abstract: Numerous deep learning-based solutions have been developed for the automatic recognition of breast cancer using mammography images. However, t

applicationsarxiv-cs-cv
14 Apr 2026
Applications

Dual-Margin Embedding for Fine-Grained Long-Tailed Plant Taxonomy

DGX agent

arXiv:2512.18994v2 Announce Type: replace Abstract: Taxonomic classification of ecological families, genera, and species underpins biodiversity monitoring and conservation. Existing computer vision me

applicationsarxiv-cs-cv
14 Apr 2026
Agents

EDFNet: Early Fusion of Edge and Depth for Thin-Obstacle Segmentation in UAV Navigation

DGX agent

arXiv:2604.09694v1 Announce Type: new Abstract: Autonomous Unmanned Aerial Vehicles (UAVs) must reliably detect thin obstacles such as wires, poles, and branches to navigate safely in real-world envir

agentsarxiv-cs-cv
14 Apr 2026
Safety

Empowering Video Translation using Multimodal Large Language Models

DGX agent

arXiv:2604.11283v1 Announce Type: new Abstract: Recent developments in video translation have further enhanced cross-lingual access to video content, with multimodal large language models (MLLMs) play

safetyarxiv-cs-cv
14 Apr 2026
Agents

Evaluating Cooperation in LLM Social Groups through Elected Leadership

DGX agent

arXiv:2604.11721v1 Announce Type: cross Abstract: Governing common-pool resources requires agents to develop enduring strategies through cooperation and self-governance to avoid collective failure. Wh

agentsarxiv-cs-ai
14 Apr 2026
Safety

Examining EAP Students' AI Disclosure Intention: A Cognition-Affect-Conation Perspective

DGX agent

arXiv:2604.10991v1 Announce Type: cross Abstract: The growing use of generative artificial intelligence (AI) in academic writing has raised increasing concerns regarding transparency and academic inte

safetyarxiv-cs-ai
14 Apr 2026
Safety

Explainability and Certification of AI-Generated Educational Assessments

DGX agent

arXiv:2604.09622v1 Announce Type: cross Abstract: The rapid adoption of generative artificial intelligence (AI) in educational assessment has created new opportunities for scalable item creation, pers

safetyarxiv-cs-ai
14 Apr 2026
Safety

Exploring the impact of fairness-aware criteria in AutoML

DGX agent

arXiv:2604.10224v1 Announce Type: cross Abstract: Machine Learning (ML) systems are increasingly used to support decision-making processes that affect individuals. However, these systems often rely on

safetyarxiv-cs-ai
14 Apr 2026
Tutorials

Find Your Optimal Teacher: Personalized Data Synthesis via Router-Guided Multi-Teacher Distillation

DGX agent

arXiv:2510.10925v2 Announce Type: replace-cross Abstract: Training student models on synthetic data generated by strong teacher models is a promising way to distilling the capabilities of teachers. Ho

tutorialsarxiv-cs-cl
14 Apr 2026
Model Releases

From GPT-3 to GPT-5: Mapping their capabilities, scope, limitations, and consequences

DGX agent

arXiv:2604.10332v1 Announce Type: new Abstract: We present the progress of the GPT family from GPT-3 through GPT-3.5, GPT-4, GPT-4 Turbo, GPT-4o, GPT-4.1, and the GPT-5 family. Our work is comparative

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

From Helpful to Trustworthy: LLM Agents for Pair Programming

DGX agent

arXiv:2604.10300v1 Announce Type: cross Abstract: LLM-based coding agents are increasingly used to generate code, tests, and documentation. Still, their outputs can be plausible yet misaligned with de

agentsarxiv-cs-ai
14 Apr 2026
Agents

GameplayQA: A Benchmarking Framework for Decision-Dense POV-Synced Multi-Video Understanding of 3D Virtual Agents

DGX agent

arXiv:2603.24329v2 Announce Type: replace-cross Abstract: Multimodal LLMs are increasingly deployed as perceptual backbones for autonomous agents in 3D environments, from robotics to virtual worlds. T

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

General-purpose LLMs as Models of Human Driver Behavior: The Case of Simplified Merging

DGX agent

arXiv:2604.09609v1 Announce Type: new Abstract: Human behavior models are essential as behavior references and for simulating human agents in virtual safety assessment of automated vehicles (AVs), yet

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

GoT-R1: Unleashing Reasoning Capability of MLLM for Visual Generation with Reinforcement Learning

DGX agent

arXiv:2505.17022v2 Announce Type: replace-cross Abstract: Visual generation models have made remarkable progress in creating realistic images from text prompts, yet struggle with complex prompts that

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

How Controllable Are Large Language Models? A Unified Evaluation across Behavioral Granularities

DGX agent

arXiv:2603.02578v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in socially sensitive domains, yet their unpredictable behaviors, ranging from misalign

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Influencing Humans to Conform to Preference Models for RLHF

DGX agent

arXiv:2501.06416v3 Announce Type: replace-cross Abstract: Designing a reinforcement learning from human feedback (RLHF) algorithm to approximate a human's unobservable reward function requires assumin

safetyarxiv-cs-ai
14 Apr 2026
Applications

Integrating SAINT with Tree-Based Models: A Case Study in Employee Attrition Prediction

DGX agent

arXiv:2604.10337v1 Announce Type: new Abstract: Employee attrition presents a major challenge for organizations, increasing costs and reducing productivity. Predicting attrition accurately enables pro

applicationsarxiv-cs-lg
14 Apr 2026
← Previous
1…452453454455456…458
Next →