AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
Local Ai

Intelligence per Watt: Measuring Intelligence Efficiency of Local AI

DGX agent

arXiv:2511.07885v4 Announce Type: replace-cross Abstract: Large language model (LLM) queries are predominantly processed by frontier models in centralized cloud infrastructure. Demand growth strains t

local-aiarxiv-cs-cl
22 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

InteractScience: Programmatic and Visually-Grounded Evaluation of Interactive Scientific Demonstration Code Generation

DGX agent

arXiv:2510.09724v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly capable of generating complete applications from natural language instructions, creating new opp

model-releasesarxiv-cs-ai
22 May 2026
Research

Internal narratives parameterise affective states

DGX agent

arXiv:2502.09487v3 Announce Type: replace Abstract: Characterising how we verbalise our feelings is central to psychological assessment and intervention, yet the mapping between narrative and affectiv

researcharxiv-cs-cl
22 May 2026
Research

Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow

DGX agent

arXiv:2605.21980v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) represent a significant leap towards empathetic agents, demonstrating remarkable capabilities in emotion understand

researcharxiv-cs-cv
22 May 2026
Model Releases

Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements

DGX agent

arXiv:2605.22079v1 Announce Type: new Abstract: Large language models (LLMs) are widely used to generate structured outputs such as JSON, SQL, and code, yet public resources remain limited for evaluat

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

JMed48k: A Multi-Profession Japanese Medical Licensing Benchmark for Vision-Language Model Evaluation

DGX agent

arXiv:2605.22080v1 Announce Type: new Abstract: We introduce JMed48k, a multi-profession Japanese healthcare licensing benchmark for evaluating vision-language models. Built from official PDF material

model-releasesarxiv-cs-cv
22 May 2026
Tutorials

La representacion de la variacion contextual mediante definiciones terminologicas flexibles

DGX agent

arXiv:1607.06330v2 Announce Type: replace Abstract: In this doctoral thesis, we apply premises of cognitive linguistics to terminological definitions and present a proposal called the flexible termino

tutorialsarxiv-cs-cl
22 May 2026
Research

Label tree semantic losses for rich multi-class medical image segmentation

DGX agent

arXiv:2507.15777v3 Announce Type: replace Abstract: Rich and accurate medical image segmentation is poised to underpin the next generation of AI-defined clinical practice by delineating critical anato

researcharxiv-cs-cv
22 May 2026
Safety

LACO: Adaptive Latent Communication for Collaborative Driving

DGX agent

arXiv:2605.22504v1 Announce Type: cross Abstract: Collaborative driving aims to improve safety and efficiency by enabling connected vehicles to coordinate under partial observability. Recent approache

safetyarxiv-cs-cv
22 May 2026
Safety

LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance

DGX agent

arXiv:2605.22567v1 Announce Type: new Abstract: Reinforcement learning has proven effective for enhancing multi-step reasoning in large language models (LLMs), yet its benefits have not fully translat

safetyarxiv-cs-cl
22 May 2026
Research

LatentOmni: Rethinking Omni-Modal Understanding via Unified Audio-Visual Latent Reasoning

DGX agent

arXiv:2605.22012v1 Announce Type: new Abstract: Joint audio-visual reasoning is essential for omnimodal understanding, yet current multimodal large language models (MLLMs) still struggle when reasonin

researcharxiv-cs-cl
22 May 2026
Agents

Learning A Unified Risk Map for Autonomous Driving in Partially Observable Environments

DGX agent

arXiv:2605.22189v1 Announce Type: new Abstract: Occlusion-aware prediction remains a critical challenge in autonomous driving due to the inherent uncertainty of unobserved regions. Existing approaches

agentsarxiv-cs-ro
22 May 2026
Safety

Learning Altruistic Collaboration in Heterogeneous Multi-Team Systems

DGX agent

arXiv:2605.21723v1 Announce Type: new Abstract: This paper studies heterogeneous multi-team collaboration through dynamic robot allocation, where robots are treated as transferable resources. Leveragi

safetyarxiv-cs-ro
22 May 2026
Model Releases

Learning Emergent Modular Representations in Multi-modality Medical Vision Foundation Models

DGX agent

arXiv:2605.21861v1 Announce Type: new Abstract: Multi-modality medical vision (MV) foundation models (FM) are fundamentally challenged by pronounced Non-IID feature statistics across heterogeneous ima

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning

DGX agent

arXiv:2605.21988v1 Announce Type: new Abstract: Video large language models (Video LLMs) achieve strong benchmark accuracy, yet often answer video questions through shortcuts such as single-frame cues

model-releasesarxiv-cs-cv
22 May 2026
Safety

Learning to Configure Agentic AI Systems

DGX agent

arXiv:2602.11574v3 Announce Type: replace Abstract: Configuring LLM-based agent systems involves choosing workflows, tools, token budgets, and prompts from a large combinatorial design space, and is t

safetyarxiv-cs-ai
22 May 2026
Safety

Learning to Evolve: Multi-modal Interactive Fields for Robust Humanoid Navigation in Dynamic Environments

DGX agent

arXiv:2605.21935v1 Announce Type: new Abstract: Safe manipulation-oriented navigation for humanoid robots requires scene memory that remains reliable under locomotion-induced perceptual distortion, en

safetyarxiv-cs-ro
22 May 2026
Model Releases

Lens: Rethinking Training Efficiency for Foundational Text-to-Image Models

DGX agent

arXiv:2605.21573v1 Announce Type: new Abstract: We introduce Lens, a 3.8B-parameter T2I model that achieves performance competitive with, and in several cases surpassing, state-of-the-art models with

model-releasesarxiv-cs-cv
22 May 2026
Research

LFX: Towards Unified Light Field Dense Semantic Segmentation and Salient Object Detection

DGX agent

arXiv:2503.00747v2 Announce Type: replace Abstract: Light field cameras capture multi-view observations within a single exposure. However, existing studies are typically tailored to specific LF repres

researcharxiv-cs-cv
22 May 2026
Agents

LIDSA: Cognitive Arbitration for Signal-Free Autonomous Intersection Management

DGX agent

arXiv:2605.12321v2 Announce Type: replace Abstract: Large language models (LLMs) show strong potential for Intelligent Transportation Systems (ITS), particularly in tasks requiring situational reasoni

agentsarxiv-cs-ai
22 May 2026
Research

LightReasoner: Can Small Language Models Teach Large Language Models Reasoning?

DGX agent

arXiv:2510.07962v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated remarkable progress in reasoning, often through supervised fine-tuning (SFT). However, SFT is resourc

researcharxiv-cs-cl
22 May 2026
Model Releases

Linear Dynamics in the RLVR Training of Large Language Models

DGX agent

arXiv:2601.04537v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven significant performance gains in reasoning-oriented large language models (LL

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

LLM Readiness Harness: Evaluation, Observability, and CI Gates for LLM/RAG Applications

DGX agent

arXiv:2603.27355v2 Announce Type: replace-cross Abstract: We present a readiness harness for LLM and RAG applications that turns evaluation into a deployment decision workflow. The system combines aut

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

LongVT: Incentivizing 'Thinking with Long Videos' via Native Tool Calling

DGX agent

arXiv:2511.20785v3 Announce Type: replace Abstract: Large multimodal models (LMMs) have shown great potential for video reasoning with textual Chain-of-Thought. However, they remain vulnerable to hall

model-releasesarxiv-cs-cv
22 May 2026
Local Ai

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming

DGX agent

arXiv:2605.21652v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have significantly advanced medical visual question answering, yet their performance in ultrasound remains suboptimal. In

local-aiarxiv-cs-cv
22 May 2026
Agents

Lower Bounds for Advection-Diffusion Equations: An Exploration with AI-Generated Proofs

DGX agent

arXiv:2605.20623v1 Announce Type: cross Abstract: We establish explicit lower bounds for advection-diffusion equations in three settings: a polynomial ot H^{-1} bound for inviscid shears with uin L^in

agentsarxiv-cs-ai
22 May 2026
Model Releases

LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model

DGX agent

arXiv:2605.22089v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising framework for end-to-end autonomous driving. However, existing VLAs typically rely on sp

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

M3: Conversational LLMs Simplify Secure Clinical Data Access, Understanding, and Analysis

DGX agent

arXiv:2507.01053v4 Announce Type: replace-cross Abstract: Large-scale clinical databases offer opportunities for medical research, but their complexity creates barriers to effective use. The Medical I

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles

DGX agent

arXiv:2605.22177v1 Announce Type: cross Abstract: The proliferation of large language models (LLMs) and modular skills has endowed autonomous agents with increasingly powerful capabilities. Existing f

model-releasesarxiv-cs-cl
22 May 2026
Tutorials

MagicFuse: Single Image Fusion for Visual and Semantic Reinforcement

DGX agent

arXiv:2602.01760v2 Announce Type: replace Abstract: This paper focuses on a highly practical scenario: how to continue benefiting from the advantages of multi-modal image fusion under harsh conditions

tutorialsarxiv-cs-cv
22 May 2026
Safety

Making the Discrete Continuous: Synthetic RAW Augmentations for Fine-Grained Evaluation of Person Detection Performance in Low Light

DGX agent

arXiv:2605.22455v1 Announce Type: new Abstract: Real-world deployment of AI vision models is both fueled and limited by the data available for training and testing. Real datasets are sparse and uneven

safetyarxiv-cs-cv
22 May 2026
Model Releases

MAP4TS: A Multi-Aspect Prompting Framework for Time-Series Forecasting with Large Language Models

DGX agent

arXiv:2510.23090v2 Announce Type: replace Abstract: Recent advances have investigated the use of pretrained large language models (LLMs) for time-series forecasting by aligning numerical inputs with L

model-releasesarxiv-cs-cl
22 May 2026
Safety

Mapping Tomato Cropping Systems in California Using AlphaEarth Geospatial Embeddings and Deep Learning Analysis

DGX agent

arXiv:2605.21804v1 Announce Type: cross Abstract: Field-scale crop maps support supply-chain forecasting and policy, yet statewide crop identification still often depends on retrospective surveys or r

safetyarxiv-cs-cv
22 May 2026
Agents

MARS: Modular Agent with Reflective Search for Automated AI Research

DGX agent

arXiv:2602.02660v3 Announce Type: replace Abstract: A critical bottleneck in automating AI research is the execution of complex machine learning engineering (MLE) tasks. MLE differs from general softw

agentsarxiv-cs-ai
22 May 2026
Model Releases

MaSC: A Masked Similarity Metric for Evaluating Concept-Driven Generation

DGX agent

arXiv:2605.22469v1 Announce Type: new Abstract: Evaluating single-concept personalization in text-to-image diffusion requires measuring both concept preservation, which captures identity fidelity to a

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Matching with Deliberation: Test-Time Evolutionary Hierarchical Multi-Agents for Zero-Shot Compositional Image Retrieval

DGX agent

arXiv:2605.22478v1 Announce Type: new Abstract: Zero-Shot Compositional Image Retrieval (ZS-CIR) requires both preserving the visual continuity of the reference image and faithfully executing the sema

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks

DGX agent

arXiv:2605.21917v1 Announce Type: new Abstract: Training Vision Language Models (VLMs) for video event reasoning requires high-quality structured annotations capturing not only what happened, but when

model-releasesarxiv-cs-cv
22 May 2026
Research

Mind the Gaps: Multi-Robot Feedback-Driven Ergodic Coverage in Unknown Environments

DGX agent

arXiv:2605.21719v1 Announce Type: new Abstract: In this work, we address the problem of multi-robot adaptive coverage, where teams of robots perform dynamic sampling by continuously adjusting their po

researcharxiv-cs-ro
22 May 2026
Model Releases

MLLMs Know When Before Speaking: Revealing and Recovering Temporal Grounding via Attention Cues

DGX agent

arXiv:2605.21954v1 Announce Type: new Abstract: Video temporal grounding (VTG), which localizes the start and end times of a queried event in an untrimmed video, is a key test of whether multimodal la

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

MM-Conv: A Multimodal Dataset and Benchmark for Context-Aware Grounding in 3D Dialogue

DGX agent

arXiv:2605.21796v1 Announce Type: cross Abstract: Grounding language in the physical world requires AI systems to interpret references that emerge dynamically during conversation. While current vision

model-releasesarxiv-cs-cl
22 May 2026
Safety

Modeling Emotional Dynamics in Agent-to-Agent Interactions on Moltbook

DGX agent

arXiv:2605.20442v1 Announce Type: cross Abstract: Generative AI systems are increasingly deployed as interactive agents in online environments, such as a social network called Moltbook. In Moltbook, l

safetyarxiv-cs-ai
22 May 2026
Safety

Modeling Pathology-Like Behavioral Patterns in Language Models Through Behavioral Fine-Tuning

DGX agent

arXiv:2605.22356v1 Announce Type: new Abstract: Large language models are increasingly used as computational tools for modeling human-like behavior. We introduce a behavioral induction framework that

safetyarxiv-cs-cl
22 May 2026
Applications

Moment-Reenacting: Inverse Motion Degradation with Cross-shutter Guidance

DGX agent

arXiv:2605.22423v1 Announce Type: new Abstract: Motion degradation, manifested as blur in global shutter (GS) images or rolling shutter (RS) distortion in RS counterparts, remains a fundamental challe

applicationsarxiv-cs-cv
22 May 2026
Safety

Moral Semantics Survive Machine Translation: Cross-Lingual Evidence from Moral Foundations Corpora

DGX agent

arXiv:2605.22660v1 Announce Type: new Abstract: Moral language is subtle and culturally variable, making it difficult to translate faithfully across languages. Idiomatic expressions, slang, and cultur

safetyarxiv-cs-cl
22 May 2026
Research

More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts

DGX agent

arXiv:2605.22641v1 Announce Type: new Abstract: Detecting Schwartz values in political text is difficult because implicit cues often depend on surrounding arguments and fine-grained distinctions betwe

researcharxiv-cs-cl
22 May 2026
Applications

MoSA: Motion-constrained Stress Adaptation for Mitigating Real-to-Sim Gap in Continuum Dynamics via Learning Residual Anisotropy

DGX agent

arXiv:2605.22597v1 Announce Type: cross Abstract: Learning real-world dynamics from visual observations is crucial for various domains. A common strategy is to calibrate simulators by estimating physi

applicationsarxiv-cs-ro
22 May 2026
Model Releases

MotiMotion: Motion-Controlled Video Generation with Visual Reasoning

DGX agent

arXiv:2605.22818v1 Announce Type: new Abstract: Current motion-controlled image-to-video generation models rigidly follow user-provided trajectories that are often sparse, imprecise, and causally inco

model-releasesarxiv-cs-cv
22 May 2026
Research

Motion Design for Grasp-Based Dynamic Locomotion in Microgravity

DGX agent

arXiv:2605.21704v1 Announce Type: new Abstract: Locomotion in microgravity often relies on sparsely and irregularly arranged anchors, motivating grasp-based mobility with multiple limbs. In this setti

researcharxiv-cs-ro
22 May 2026
← Previous
1…794795796797798…1311
Next →