AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,134 results
Agents

Building AI-Ready Data Systems for Space Life Sciences, Aerospace Medicine, and Deep Space Exploration

DGX agent

arXiv:2606.28856v1 Announce Type: cross Abstract: While AI holds the potential to revolutionize space life sciences, realizing this promise is contingent upon the systematic restructuring of heterogen

agentsarxiv-cs-ai
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested

DGX agent

arXiv:2606.28430v1 Announce Type: cross Abstract: Benchmarks are widely used to evaluate task completion by Large Language Models (LLMs), but this approach has accumulated construction-validity proble

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Causality for Tabular Data Synthesis: A High-Order Structure Causal Benchmark Framework

DGX agent

arXiv:2406.08311v3 Announce Type: replace-cross Abstract: Existing evaluations of tabular synthesis models rely primarily on low-order statistics and downstream task performance, leaving multivariate

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

CaveAgent: Transforming LLMs into Stateful Runtime Operators

DGX agent

arXiv:2601.01569v4 Announce Type: replace Abstract: LLM-based agents are increasingly capable of complex task execution, yet current agentic systems remain constrained by text-centric paradigms that s

agentsarxiv-cs-ai
30 Jun 2026
Agents

Code Reasoning for Software Engineering Tasks: A Survey and A Call to Action

DGX agent

arXiv:2506.13932v3 Announce Type: replace-cross Abstract: The rise of large language models (LLMs) has led to dramatic improvements across a wide range of natural language tasks. Their performance on

agentsarxiv-cs-ai
30 Jun 2026
Applications

Comparing Human and Automatic Recognition of Dutch Dysarthric Continuous Speech: A Case Study

DGX agent

arXiv:2606.30237v1 Announce Type: new Abstract: In our goal to develop personalised dysarthric speech recognition (DSR) models, this study compared the recognition performances of human listeners and

applicationsarxiv-cs-cl
30 Jun 2026
Model Releases

Customized Generative AI Agent for Transportation Engineering Practice: A Development and Continued Pre-training Guideline

DGX agent

arXiv:2606.29014v1 Announce Type: new Abstract: Recent advancements in generative artificial intelligence (AI) and large language models (LLMs) have shown significant promise in automating complex rea

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Defeat Devices in AI Systems

DGX agent

arXiv:2606.28863v1 Announce Type: cross Abstract: AI systems increasingly exhibit behavior that differs systematically between evaluation and deployment contexts. Alignment faking, sandbagging, benchm

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Direct Causation in International Humanitarian Law and the Challenge of AI-Mediated Civilian Cyber Operations

DGX agent

arXiv:2606.29175v1 Announce Type: new Abstract: International humanitarian law protects civilians from direct attack unless and for such time as they take direct part in hostilities, with the ICRC's 2

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

DiscoGen: Procedural Generation of Algorithm Discovery Tasks in Machine Learning

DGX agent

arXiv:2603.17863v2 Announce Type: replace-cross Abstract: Automating the development of machine learning algorithms has the potential to unlock new breakthroughs. However, our ability to improve and e

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks

DGX agent

arXiv:2510.14207v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are powering a growing share of interactive web applications, yet remain vulnerable to misuse and harm. Prior jail

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

EPIC-EuroParl-UdS: Information-Theoretic Perspectives on Translation and Interpreting

DGX agent

arXiv:2603.09785v3 Announce Type: replace Abstract: This paper introduces an updated and combined version of the bidirectional English-German EPIC-UdS (spoken) and EuroParl-UdS (written) corpora conta

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions

DGX agent

arXiv:2507.05257v4 Announce Type: replace-cross Abstract: Recent benchmarks for Large Language Model (LLM) agents primarily focus on evaluating reasoning, planning, and execution capabilities, while a

model-releasesarxiv-cs-ai
30 Jun 2026
Applications

Exploring the Value of Diverse LLM Explanations in Introductory Programming

DGX agent

arXiv:2606.28882v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown the potential to generate code explanations that surpass those of peers in quality, offering promising opportu

applicationsarxiv-cs-ai
30 Jun 2026
Model Releases

fev-bench: A Realistic Benchmark for Time Series Forecasting

DGX agent

arXiv:2509.26468v3 Announce Type: replace Abstract: Benchmark quality is critical for meaningful evaluation and sustained progress in time series forecasting, particularly with the rise of pretrained

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

From Failure Taxonomy to Intervention: A Diagnostic Methodology for Industry-Scale AVLM in Video and Live-Streaming Platform Moderation

DGX agent

arXiv:2606.30059v1 Announce Type: new Abstract: Industry-scale video and live-streaming moderation imposes requirements that are difficult to satisfy with generic pretrained public models or external

model-releasesarxiv-cs-lg
30 Jun 2026
Safety

FutureNav: Unified World-Action Modeling for Vision-and-Language Navigation

DGX agent

arXiv:2606.30367v1 Announce Type: new Abstract: Vision-and-language navigation (VLN) in continuous environments requires an agent to ground instructions in egocentric observations while maintaining sp

safetyarxiv-cs-ro
30 Jun 2026
Applications

GaRLILEO: Gravity-aligned Radar-Leg-Inertial Enhanced Odometry

DGX agent

arXiv:2511.13216v2 Announce Type: replace Abstract: Deployment of legged robots for navigating challenging terrains (e.g., stairs, slopes, and unstructured environments) has gained increasing preferen

applicationsarxiv-cs-ro
30 Jun 2026
Safety

Generative Learning as a Tool to Improve Perception of Emotional Body Motion Expressions

DGX agent

arXiv:2606.28769v1 Announce Type: new Abstract: Emotional body motion expressions are an essential element of non-verbal communication. Effectively conveying these expressions through technology is of

safetyarxiv-cs-lg
30 Jun 2026
Safety

GeoISF: Instance Semantic Forest Inspired Large-Scale Cross-View Geo-Localization via Ground LiDAR-to-Satellite Image

DGX agent

arXiv:2606.28371v1 Announce Type: new Abstract: The problem of localization on a large-scale satellite image given a frame of query ground view point clouds remains challenging. Existing LiDAR-to-imag

safetyarxiv-cs-cv
30 Jun 2026
Model Releases

Grounding Sim-to-Real Generalization in Robotic Manipulation: An Empirical Study with Vision-Language-Action Models

DGX agent

arXiv:2603.22876v2 Announce Type: replace-cross Abstract: Learning a generalist control policy for robotic manipulation typically relies on large-scale datasets. Given the high cost of real-world data

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

How Far Can You Get Without a GPU? A Systematic Benchmark of Lightweight Hallucination Detection Across Question Answering, Dialogue, and Summarisation

DGX agent

arXiv:2606.29809v1 Announce Type: cross Abstract: Hallucination detection has become a pressing requirement for trustworthy AI deployment at scale. The most accurate detection methods depend on GPU-in

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

How LLMs See Creativity: Zero-Shot Scoring of Visual Creativity with Interpretable Reasoning

DGX agent

arXiv:2606.29672v1 Announce Type: new Abstract: Evaluating the originality of visual images poses enduring challenges for creativity assessment. Automated scoring using AI models has proven effective

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

InsertAnywhere: Geometrically Grounded and Optics-Aware Video Object Insertion

DGX agent

arXiv:2512.17504v2 Announce Type: replace-cross Abstract: Recent advances in diffusion models have enabled impressive video editing capabilities, yet production-grade Video Object Insertion (VOI) rema

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Learning from Reliable Latent Prompts for Visual Recognition with Missing Modalities

DGX agent

arXiv:2606.30597v1 Announce Type: new Abstract: Large-scale multimodal models (LMMs) have achieved superior performance in visual recognition by synergizing information across diverse, massive-scale p

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Learning to Segment Liquids in Real-world Images

DGX agent

arXiv:2601.00940v2 Announce Type: replace Abstract: Liquids like water, wine and medicine are everywhere. However, limited attention has been given to the task of segmenting liquids, hindering the abi

model-releasesarxiv-cs-cv
30 Jun 2026
Local Ai

Learning Where and When: Patch-Based Spatiotemporal Localization in Weakly Supervised Video Anomaly Detection

DGX agent

arXiv:2606.29498v1 Announce Type: new Abstract: Weakly supervised video anomaly detection (WSVAD) has predominantly focused on temporal localization, identifying when anomalies occur while largely neg

local-aiarxiv-cs-cv
30 Jun 2026
Agents

LLM agents security duality: a comprehensive survey of self-security and empowered cybersecurity

DGX agent

arXiv:2606.28450v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly being integrated into real-world systems. Their autonomy and tool-use capabilities generate substantial

agentsarxiv-cs-ai
30 Jun 2026
Local Ai

Low-cost concept-based localized explanations: How far can we get with training-free approaches?

DGX agent

arXiv:2606.29069v1 Announce Type: new Abstract: Concept-based Explainable AI (C-XAI) seeks human-understandable explanations grounded in semantic concepts, yet validation is limited by the scarcity of

local-aiarxiv-cs-ai
30 Jun 2026
Model Releases

MaDI-Bench: An End-to-End Data Integration Benchmark

DGX agent

arXiv:2606.30371v1 Announce Type: cross Abstract: Data integration combines heterogeneous data sets into a single, coherent representation. Data integration involves a sequence of interdependent tasks

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

MAM-AI: An On-Device Medical Retrieval-Augmented Generation System for Nurses and Midwives in Zanzibar

DGX agent

arXiv:2606.29580v1 Announce Type: new Abstract: Maternal and newborn mortality remain among the highest in sub-Saharan Africa, where midwifery care is often delivered by nurses who lack midwifery trai

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

mamabench and mamaretrieval: Benchmarks for Evaluating Medical Retrieval-Augmented Generation in Maternal, Neonatal, and Reproductive Health

DGX agent

arXiv:2606.29467v1 Announce Type: new Abstract: Medical question-answering benchmarks rarely cover the maternal, neonatal, child, and reproductive-health questions a nurse-midwife asks, and, to our kn

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

MARS: A neurosymbolic approach for interpretable drug discovery

DGX agent

arXiv:2410.05289v4 Announce Type: replace Abstract: Background: Neurosymbolic (NeSy) artificial intelligence describes the combination of logic or rule-based techniques with neural networks. Compared

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Memory-Managed Long-Context Attention: A Preliminary Study of Editable Request-Local Memory

DGX agent

arXiv:2606.28876v1 Announce Type: new Abstract: Long-context language models often conflate two different goals: compressing history into an efficient state, and maintaining reliable long-term memory.

model-releasesarxiv-cs-cl
30 Jun 2026
Agents

Modeling Earth-Scale Human-Like Societies with One Billion Agents

DGX agent

arXiv:2506.12078v2 Announce Type: replace-cross Abstract: Understanding the dynamic evolution of complex social phenomena requires both high-fidelity modeling of human behavior and large-scale simulat

agentsarxiv-cs-ai
30 Jun 2026
Agents

MonoSR: Open-Vocabulary Spatial Reasoning from Monocular Images

DGX agent

arXiv:2511.19119v2 Announce Type: replace Abstract: Spatial reasoning (SR), the ability to infer 3D spatial information from 2D inputs, is essential for real-world applications such as embodied AI and

agentsarxiv-cs-cv
30 Jun 2026
Safety

Multi-Class Human/Object Detection on Robot Manipulators using Proprioceptive Sensing

DGX agent

arXiv:2508.02425v2 Announce Type: replace-cross Abstract: In physical human-robot collaboration (pHRC) settings, humans and robots collaborate directly in shared environments. Robots must analyze inte

safetyarxiv-cs-ai
30 Jun 2026
Safety

Multimodal Representation Alignment for Cross-modal Information Retrieval

DGX agent

arXiv:2506.08774v2 Announce Type: replace-cross Abstract: Different machine learning models can represent the same underlying concept in different ways. This variability is particularly valuable for i

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Not-quite-human tastes: the stylized omnivorousness of LLM survey surrogates

DGX agent

arXiv:2606.30085v1 Announce Type: new Abstract: Large-language models have proven to be remarkable if inconsistent parrots of public attitudes and opinions. The extent to which LLMs are able to produc

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

OmniCoT: A Benchmark for Global and Multi-Step Panoramic Reasoning

DGX agent

arXiv:2606.30378v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated promising spatial reasoning capabilities, while these abilities remain underexplored in the e

model-releasesarxiv-cs-cv
30 Jun 2026
Local Ai

Online Data Selection for Instruction Tuning via Gaussian Processes

DGX agent

arXiv:2606.30077v1 Announce Type: cross Abstract: With Large Language Model (LLM) pre-training and fine-tuning shifting its focus from data volume to data quality, quality data selection has emerged a

local-aiarxiv-cs-ai
30 Jun 2026
Safety

Persona-Trained Monte Carlo: Estimating Market-Outcome Distributions via Swarms of Persona-Conditioned Neural Policy Bots in a Limit Order Book

DGX agent

arXiv:2606.29556v1 Announce Type: new Abstract: We propose Persona-Trained Monte Carlo (PTMC), a method for estimating distributions of market-outcome statistics by repeatedly simulating limit-order-b

safetyarxiv-cs-lg
30 Jun 2026
Agents

Reinforcement Learning for Software Vulnerability Analysis: A Systematic Review with Emphasis on C/C++ Source Code and Static Analysis

DGX agent

arXiv:2606.28403v1 Announce Type: cross Abstract: Vulnerability detection in C/C++ software remains a major security challenge due to code complexity, manual memory management, and the limitations of

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation

DGX agent

arXiv:2606.28998v1 Announce Type: cross Abstract: Large Language Model (LLM) alignment trains an LLM using preference data to produce outputs that better meet established quality standards. While LLM

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation

DGX agent

arXiv:2602.09305v2 Announce Type: replace Abstract: Large Language Models (LLMs) demonstrate transformative potential, yet their reasoning remains inconsistent and unreliable. Reinforcement learning (

safetyarxiv-cs-lg
30 Jun 2026
Model Releases

SA-Homo: Scale Adaptive Homography Estimation for Scale Variation Scenarios

DGX agent

arXiv:2606.30408v1 Announce Type: new Abstract: Homography estimation, as one of the fundamental problems in computer vision, remains challenged by scale variation scenarios where image pairs potentia

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SADL: What to Ignore? A Benchmark for Subject-Aware Distractor Localization

DGX agent

arXiv:2606.30393v1 Announce Type: new Abstract: Photographs frequently contain visual distractors besides foregrounds and backgrounds of the intended subject, competing for attention and weakening com

model-releasesarxiv-cs-cv
30 Jun 2026
Applications

Solver-Verified Formulation Generation and Selection for Multi-Warehouse Inventory Allocation Using Large Language Models

DGX agent

arXiv:2606.29366v1 Announce Type: cross Abstract: Balance-oriented multi-warehouse inventory allocation is a recurring decision problem in large-scale e-commerce supply chains, in which a fixed replen

applicationsarxiv-cs-ai
30 Jun 2026
← Previous
1…420421422423424…462
Next →