AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

How Can Driving World Models Do Counterfactual Prediction?

DGX agent

arXiv:2608.11601v1 Announce Type: new Abstract: Driving world models are often interpreted as counterfactual simulators for observed driving episodes: given a factual driving log, they are asked what

model-releasesarxiv-cs-cv
13 Aug 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Lost in Compaction: Evaluating Side-Constraint Loss under Context Compaction

DGX agent

arXiv:2608.11242v1 Announce Type: cross Abstract: When the context window is under pressure, LLM systems compact prior context to continue ongoing tasks. We identify a class of user-issued instruction

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence

DGX agent

arXiv:2608.12036v1 Announce Type: new Abstract: AI models have achieved remarkable success across diverse domains, yet the mechanisms underlying their capabilities and the risks they may pose remain p

model-releasesarxiv-cs-ai
13 Aug 2026
Applications

Methodologies for Improving the Quality of AI Tutoring in K-12 Education

DGX agent

arXiv:2608.11259v1 Announce Type: cross Abstract: Many AI tutors leverage large language models (LLMs) today. Given that LLMs are opaque black boxes, robust evaluation and live experimentation to meas

applicationsarxiv-cs-ai
13 Aug 2026
Model Releases

Physics-Informed Implicit Neural Representations for Improved Myocardial Perfusion MRI Quantification

DGX agent

arXiv:2608.11282v1 Announce Type: cross Abstract: Quantifying myocardial perfusion from cardiac magnetic resonance (CMR) can be achieved by fitting tracer-kinetic models to the dynamic contrast-enhanc

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

SAG: SQL-Retrieval Augmented Generation with Query-Time Dynamic Hyperedges

DGX agent

arXiv:2608.12129v1 Announce Type: new Abstract: While retrieval-augmented generation (RAG) has proven effective at giving LLMs access to external knowledge, mainstream dense-retrieval implementations

model-releasesarxiv-cs-cl
13 Aug 2026
Model Releases

SCOPE-Router: Cost-Aware Open-Set VLM Routing for Execution-Oriented Tasks

DGX agent

arXiv:2608.12127v1 Announce Type: new Abstract: Model routing aims to select the most suitable model from a candidate pool for each query, balancing quality and cost. Existing VLM routing research is

model-releasesarxiv-cs-cv
13 Aug 2026
Model Releases

VICBench: A Multi-Language Benchmark for Code Vulnerability Detection

DGX agent

arXiv:2608.12246v1 Announce Type: cross Abstract: Evaluating security vulnerability detection tools requires benchmark datasets with vulnerability-inducing commits (VICs) - the commits that first intr

model-releasesarxiv-cs-ai
13 Aug 2026
Safety

Video2Track: From Real-World Interaction Videos to Steerable Adversarial Closed-Track Testing for Automated Driving Systems

DGX agent

arXiv:2608.11592v1 Announce Type: new Abstract: Closed-track testing plays a fundamental role in the verification and validation of automated driving systems (ADS), particularly for safety-critical sc

safetyarxiv-cs-ro
13 Aug 2026
Model Releases

AIFS-TC: A simple correction competitive with the operational frontier for tropical cyclone intensity forecasting

DGX agent

arXiv:2608.09959v1 Announce Type: cross Abstract: AI weather models are in the process of revolutionising weather forecasting. While these models have been shown to achieve superior performance to phy

model-releasesarxiv-cs-lg
12 Aug 2026
Local Ai

Beyond Detection: Evaluating Defensive LLMs Against AI-Generated Social Engineering in Live Turn-by-Turn Interaction

DGX agent

arXiv:2608.10239v1 Announce Type: new Abstract: Generative AI makes social-engineering attacks more fluent, adaptive, and scalable, increasing the need for LLM-based de- fenders that can protect users

local-aiarxiv-cs-ai
12 Aug 2026
Research

Do LLMs Benefit From Their Own Words?

DGX agent

arXiv:2602.24287v2 Announce Type: replace-cross Abstract: In multi-turn conversations, large language models typically condition on the full conversation history: both past user prompts and assistant

researcharxiv-cs-ai
12 Aug 2026
Model Releases

DriveVLA-M0: Failure-Aware Memory Augmentation for Autonomous Driving

DGX agent

arXiv:2608.10413v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for end-to-end autonomous driving by enabling unified reasoning across

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

E^3mo-Bench: A Scalable Benchmark for Multimodal Evoked and Expressed Emotion Understanding via Bayesian Pairwise Alignment

DGX agent

arXiv:2608.10796v1 Announce Type: new Abstract: Understanding both expressed and evoked emotions is critical for multimodal large language models (MLLMs) to achieve comprehensive affect-aware interact

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

Fast and Memory-Efficient Wavelet Convolutions via I/O-Aware Reformulation

DGX agent

arXiv:2608.10805v1 Announce Type: cross Abstract: Wavelet convolution (WTConv) has emerged as an increasingly popular drop-in replacement for standard convolutions, expanding a network's receptive fie

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

From Reasoning Depth to Reasoning Breadth: Evaluating Multi-Point Associative Reasoning in Large Language Models

DGX agent

arXiv:2608.10444v1 Announce Type: cross Abstract: Large language models (LLMs) have made substantial progress on reasoning tasks that require increasingly long and complex inferential chains. This pro

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes

DGX agent

arXiv:2608.10886v1 Announce Type: new Abstract: Robots operating in human environments need memories that capture not only what objects exist and where, but also how people use them over time and how

model-releasesarxiv-cs-cv
12 Aug 2026
Safety

Hidden in Plain Sight: Diffusion-Based Unrestricted Robotic Attacks on Vision-Language-Action Models

DGX agent

arXiv:2608.10393v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong capabilities in controlling robots across diverse manipulation tasks. However, their adversarial r

safetyarxiv-cs-ai
12 Aug 2026
Model Releases

HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

DGX agent

arXiv:2506.03922v4 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated significant potential to advance a broad range of domains. However, current benchma

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

HUI360: A 360{eg} Egocentric Dataset and Baselines for Human-Robot Interaction Anticipation

DGX agent

arXiv:2608.11051v1 Announce Type: new Abstract: As robots increasingly operate in human-populated environments, anticipating human intentions is essential for enabling proactive and socially aware beh

model-releasesarxiv-cs-cv
12 Aug 2026
Safety

MedUP: Awakening Unified Understanding and Perception in Medical Vision-Language Models

DGX agent

arXiv:2608.10635v1 Announce Type: cross Abstract: Medical Vision-Language Models (Med-VLMs) excel at verbalizing visual content, yet precise visual perception, segmentation, and grounding remain chall

safetyarxiv-cs-ai
12 Aug 2026
Safety

Observational Policy Ranking for SMB Financial Guidance from Multi-Action Accounting Logs

DGX agent

arXiv:2608.10050v1 Announce Type: new Abstract: Small and medium-sized businesses need timely financial guidance, yet historical accounting logs record self-selected and often co-occurring business ch

safetyarxiv-cs-lg
12 Aug 2026
Model Releases

Optimal Stopping of Self-Refining Foundation Models

DGX agent

arXiv:2608.10729v1 Announce Type: cross Abstract: Foundation models can improve their outputs through a self-refinement process driven by external feedback. In this process, the model is embedded in a

model-releasesarxiv-cs-ai
12 Aug 2026
Research

Order Matters: LVLMs as Judges for Temporal Reasoning in Image Sequences

DGX agent

arXiv:2608.10908v1 Announce Type: cross Abstract: As generative multimedia evolves from static image synthesis to complex, interleaved visual narratives, a foundational bottleneck has emerged: the jud

researcharxiv-cs-cl
12 Aug 2026
Safety

Predicting Space Groups of Double Perovskites by LLM with Dynamic Few-Shot Learning

DGX agent

arXiv:2608.10483v1 Announce Type: new Abstract: Double perovskites (DPs) offer broad compositional tunability, but predicting the space groups (SGs) of stable structures remains difficult because avai

safetyarxiv-cs-ai
12 Aug 2026
Safety

Procedural Fairness Failures in RLHF from Preference Averaging

DGX agent

arXiv:2608.10126v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) aggregates heterogeneous preferences into a single reward model, assuming preference homogeneity. Wh

safetyarxiv-cs-ai
12 Aug 2026
Local Ai

Quantum Coordination Advantages in AI State-Tracking Tasks: Semantic Compilation and Latent Memory

DGX agent

arXiv:2608.11066v1 Announce Type: cross Abstract: We prove inference-time quantum coordination advantages for specified AI state-tracking tasks. A solver compresses semantic history into a future-acce

local-aiarxiv-cs-ai
12 Aug 2026
Local Ai

R4DSG: Relative 4D Scene Graph Memory for Object-Centric Question Answering in Long Egocentric Video

DGX agent

arXiv:2608.11017v1 Announce Type: cross Abstract: Long-horizon egocentric video is a rich substrate for wearable AI assistants, but object-centric questions such as where an item was moved, when it la

local-aiarxiv-cs-ai
12 Aug 2026
Model Releases

ReLTEx: Reliable LLM-based Taxonomy Expansion

DGX agent

arXiv:2608.10970v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have demonstrated strong capabilities in generating semantically relevant concepts and relations, maki

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Rethinking Text-Based Image Retrieval in Specific Domain

DGX agent

arXiv:2608.10524v1 Announce Type: cross Abstract: Driven by the rapid advancement of vision-language representation learning, Text-based Image Retrieval (TBIR) has made notable progress. However, exis

model-releasesarxiv-cs-ai
12 Aug 2026
Safety

Scheduling Mixed RL Rollouts Beyond Prefix Locality

DGX agent

arXiv:2608.11152v1 Announce Type: cross Abstract: Modern reinforcement learning (RL) post-training pipelines for large language models (LLMs) increasingly combine rollout workloads across multiple dom

safetyarxiv-cs-lg
12 Aug 2026
Research

StreamFlow: Dynamic Memory Flows for Streaming Video Understanding

DGX agent

arXiv:2608.10949v1 Announce Type: cross Abstract: Streaming video understanding requires multimodal large language models (MLLMs) to preserve relevant evidence from continuously evolving streams under

researcharxiv-cs-cl
12 Aug 2026
Safety

Toward a Theory of Value in AI Alignment

DGX agent

arXiv:2608.10327v1 Announce Type: new Abstract: Can AI systems be aligned to human values? The popularization of large language models (LLMs) and multi-modal foundation models has seen a rise in harms

safetyarxiv-cs-ai
12 Aug 2026
Model Releases

When Chain-of-Thought Helps and When It Hurts: An Empirical Investigation of the Serial-Depth Bottleneck in LLM Reasoning

DGX agent

arXiv:2608.09942v1 Announce Type: cross Abstract: It is widely assumed that chain-of-thought (CoT) prompting universally improves LLM reasoning. We investigate this through the conceptual framework of

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Workflow Cards: Structured Summaries of Workflow Executions Using Provenance Data

DGX agent

arXiv:2608.11022v1 Announce Type: cross Abstract: Model Cards and Data Cards have demonstrated the value of structured, human-readable documentation for machine learning artifacts, capturing their con

model-releasesarxiv-cs-ai
12 Aug 2026
Research

A foundation model of numerical intelligence with cross-disciplinary generalization

DGX agent

arXiv:2607.28432v2 Announce Type: replace Abstract: Intelligence is commonly understood as the ability to acquire and apply knowledge, adapt to unfamiliar situations and solve new problems. Large lang

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Autorubric: A Unifying Framework for Rubric-Based LLM Evaluation on Non-Verifiable Tasks

DGX agent

arXiv:2603.00077v3 Announce Type: replace-cross Abstract: Rubric-based LLM judges have become indispensable for evaluating and optimizing systems on non-verifiable tasks, where success cannot be reduc

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

ComplexityWorld: Benchmarking Vision-Language Models on Verifiable Visual Decision Making

DGX agent

arXiv:2608.07584v1 Announce Type: new Abstract: Vision-language models (VLMs) have made rapid progress in visual perception and increasingly support real-world tasks that depend on images. Many such t

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Cross-Model Humor Preference Modeling with Cards Against Humanity

DGX agent

arXiv:2608.07481v1 Announce Type: cross Abstract: This paper investigates whether one large language model can approximate the humor preferences of another in a controlled Cards Against Humanity-style

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

DH-VLM: Dual-Horizon Cooperative Latent Reasoning for Autonomous Driving

DGX agent

arXiv:2608.09333v1 Announce Type: new Abstract: Large-scale language models for autonomous driving enable enhanced global understanding and long-horizon planning. However, when deployed in isolated ve

safetyarxiv-cs-ro
11 Aug 2026
Safety

Dual-Adversarial Safety Alignment: Cultivating Intrinsic Threat Comprehension in LRMs

DGX agent

arXiv:2608.09542v1 Announce Type: cross Abstract: Large reasoning models (LRMs) achieve remarkable success on complex tasks but remain vulnerable to harmful prompts that induce unsafe outputs. Recent

safetyarxiv-cs-ai
11 Aug 2026
Safety

Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition

DGX agent

arXiv:2608.09762v1 Announce Type: new Abstract: Real-world online reinforcement learning (RL) provides a promising approach for training robotic manipulation policies directly in the physical world, a

safetyarxiv-cs-ro
11 Aug 2026
Model Releases

ELICITED: EHR-grounded Longitudinal Interactive Conversations for Information-seeking Triage Evaluation and Decision-making

DGX agent

arXiv:2608.09024v1 Announce Type: new Abstract: Emergency-department (ED) triage requires clinicians to rapidly identify patients who need immediate attention, determine who can safely wait, and prior

model-releasesarxiv-cs-cl
11 Aug 2026
Research

Emotion in an active inference model of human driving

DGX agent

arXiv:2608.07480v1 Announce Type: new Abstract: Active inference has emerged as a principled framework for modeling adaptive behavior by balancing goal-directed action with uncertainty reduction. It h

researcharxiv-cs-ai
11 Aug 2026
Research

EvalConvoLearn: An Open-Source Framework for Evaluating Grounded Learner Simulations in Tutoring Conversations

DGX agent

arXiv:2608.07497v1 Announce Type: cross Abstract: Conversational learner simulations are valuable tools for testing learning theories, evaluating instructional materials and automated tutors, or power

researcharxiv-cs-cl
11 Aug 2026
Model Releases

FaLCon: Facet-Anchored Retrieval with Late Consensus for Sim2Real Text-Based Person Anomaly Search

DGX agent

arXiv:2608.09474v1 Announce Type: new Abstract: Text-based person anomaly search requires retrieving real-world pedestrian images from detailed natural-language descriptions using models trained prima

model-releasesarxiv-cs-cv
11 Aug 2026
Safety

Forgotten History or Test-of-Time? Retrospect and Prospect on RAG from an IR Perspective

DGX agent

arXiv:2608.08445v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) is widely regarded as a novel paradigm born from the limitations of large language models (LLMs)--a mechanism to gr

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

From Diagnosis to Correction: Benchmarking and Improving Real-World Table Parsing

DGX agent

arXiv:2608.09842v1 Announce Type: new Abstract: Recent document parsers achieve table TEDS scores above 93 on OmniDocBench v1.6, yet community feedback and our audit reveal persistent failures on comp

model-releasesarxiv-cs-cv
11 Aug 2026
← Previous
1…202203204205206…233
Next →