AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
Human
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,006 results
22 May 2026

Impact of Atmospheric Turbulence and Pointing Error on Earth Observation

ApplicationsDGX agent

arXiv:2605.22268v1 Announce Type: cross Abstract: Earth Observation (EO) imagery is often degraded by atmospheric turbulence and pointing jitter; yet, these effects are rarely considered in datasets u

Improved DDIM Sampling with Moment Matching Gaussian Mixtures

ResearchDGX agent

arXiv:2311.04938v5 Announce Type: replace Abstract: We propose using a Gaussian Mixture Model (GMM) as reverse transition operator (kernel) within the Denoising Diffusion Implicit Models (DDIM) framew

ImProver: Agent-Based Automated Proof Optimization

AgentsDGX agent

arXiv:2410.04753v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been used to generate formal proofs of mathematical theorems in proofs assistants such as Lean. However, we

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Improving 3D Labeling in Self-Driving by Inferring Vehicle Information using Vision Language Models

ResearchDGX agent

arXiv:2605.21747v1 Announce Type: new Abstract: We present an approach to improve 3D vehicle labeling in self-driving applications through zero-shot inference of vehicle information, leveraging Vehicl

Improving Viewpoint-Invariance and Temporal Consistency for Action Detection

ResearchDGX agent

arXiv:2605.22695v1 Announce Type: new Abstract: Viewpoint change invariance and action temporal consistency are critical aspects for the effective deployment of human action detection of untrimmed vid

In Silico Modeling of the RAMPHO Buffer: Dissociating Informational and Energetic Masking via Phonetic Entropy in Deep Neural Networks

ResearchDGX agent

arXiv:2605.22465v1 Announce Type: new Abstract: The fundamental challenge of listening in multi-talker environments is a cognitive bottleneck, defined by the Ease of Language Understanding (ELU) model

Industrial Dual-Arm Box Handling via Online Inertial Estimation and Convex Wrench Optimization

ResearchDGX agent

arXiv:2605.22021v1 Announce Type: new Abstract: Industrial robotic object handling often involves boxes and packages whose mass and center of mass are not known in advance. These uncertainties affect

InfVSR: Breaking Length Limits of Generic Video Super-Resolution

Model ReleasesDGX agent

arXiv:2510.00948v2 Announce Type: replace Abstract: Real-world videos often extend over thousands of frames. Existing generative video super-resolution (VSR) approaches, however, face two persistent c

InnerQ: Hardware-Aware Tuning-Free Quantization of KV Cache for Large Language Models

Model ReleasesDGX agent

arXiv:2602.23200v2 Announce Type: replace-cross Abstract: When transformer-based language models are deployed for text generation, most of the inference time is spent in the decoding stage, where outp

Intelligence per Watt: Measuring Intelligence Efficiency of Local AI

Local AiDGX agent

arXiv:2511.07885v4 Announce Type: replace-cross Abstract: Large language model (LLM) queries are predominantly processed by frontier models in centralized cloud infrastructure. Demand growth strains t

InteractScience: Programmatic and Visually-Grounded Evaluation of Interactive Scientific Demonstration Code Generation

Model ReleasesDGX agent

arXiv:2510.09724v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly capable of generating complete applications from natural language instructions, creating new opp

Internal narratives parameterise affective states

ResearchDGX agent

arXiv:2502.09487v3 Announce Type: replace Abstract: Characterising how we verbalise our feelings is central to psychological assessment and intervention, yet the mapping between narrative and affectiv

Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow

ResearchDGX agent

arXiv:2605.21980v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) represent a significant leap towards empathetic agents, demonstrating remarkable capabilities in emotion understand

Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements

Model ReleasesDGX agent

arXiv:2605.22079v1 Announce Type: new Abstract: Large language models (LLMs) are widely used to generate structured outputs such as JSON, SQL, and code, yet public resources remain limited for evaluat

JMed48k: A Multi-Profession Japanese Medical Licensing Benchmark for Vision-Language Model Evaluation

Model ReleasesDGX agent

arXiv:2605.22080v1 Announce Type: new Abstract: We introduce JMed48k, a multi-profession Japanese healthcare licensing benchmark for evaluating vision-language models. Built from official PDF material

La representacion de la variacion contextual mediante definiciones terminologicas flexibles

TutorialsDGX agent

arXiv:1607.06330v2 Announce Type: replace Abstract: In this doctoral thesis, we apply premises of cognitive linguistics to terminological definitions and present a proposal called the flexible termino

Label tree semantic losses for rich multi-class medical image segmentation

ResearchDGX agent

arXiv:2507.15777v3 Announce Type: replace Abstract: Rich and accurate medical image segmentation is poised to underpin the next generation of AI-defined clinical practice by delineating critical anato

LACO: Adaptive Latent Communication for Collaborative Driving

SafetyDGX agent

arXiv:2605.22504v1 Announce Type: cross Abstract: Collaborative driving aims to improve safety and efficiency by enabling connected vehicles to coordinate under partial observability. Recent approache

LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance

SafetyDGX agent

arXiv:2605.22567v1 Announce Type: new Abstract: Reinforcement learning has proven effective for enhancing multi-step reasoning in large language models (LLMs), yet its benefits have not fully translat

LatentOmni: Rethinking Omni-Modal Understanding via Unified Audio-Visual Latent Reasoning

ResearchDGX agent

arXiv:2605.22012v1 Announce Type: new Abstract: Joint audio-visual reasoning is essential for omnimodal understanding, yet current multimodal large language models (MLLMs) still struggle when reasonin

Learning A Unified Risk Map for Autonomous Driving in Partially Observable Environments

AgentsDGX agent

arXiv:2605.22189v1 Announce Type: new Abstract: Occlusion-aware prediction remains a critical challenge in autonomous driving due to the inherent uncertainty of unobserved regions. Existing approaches

Learning Altruistic Collaboration in Heterogeneous Multi-Team Systems

SafetyDGX agent

arXiv:2605.21723v1 Announce Type: new Abstract: This paper studies heterogeneous multi-team collaboration through dynamic robot allocation, where robots are treated as transferable resources. Leveragi

Learning Emergent Modular Representations in Multi-modality Medical Vision Foundation Models

Model ReleasesDGX agent

arXiv:2605.21861v1 Announce Type: new Abstract: Multi-modality medical vision (MV) foundation models (FM) are fundamentally challenged by pronounced Non-IID feature statistics across heterogeneous ima

Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.21988v1 Announce Type: new Abstract: Video large language models (Video LLMs) achieve strong benchmark accuracy, yet often answer video questions through shortcuts such as single-frame cues

Learning to Configure Agentic AI Systems

SafetyDGX agent

arXiv:2602.11574v3 Announce Type: replace Abstract: Configuring LLM-based agent systems involves choosing workflows, tools, token budgets, and prompts from a large combinatorial design space, and is t

Learning to Evolve: Multi-modal Interactive Fields for Robust Humanoid Navigation in Dynamic Environments

SafetyDGX agent

arXiv:2605.21935v1 Announce Type: new Abstract: Safe manipulation-oriented navigation for humanoid robots requires scene memory that remains reliable under locomotion-induced perceptual distortion, en

Lens: Rethinking Training Efficiency for Foundational Text-to-Image Models

Model ReleasesDGX agent

arXiv:2605.21573v1 Announce Type: new Abstract: We introduce Lens, a 3.8B-parameter T2I model that achieves performance competitive with, and in several cases surpassing, state-of-the-art models with

LFX: Towards Unified Light Field Dense Semantic Segmentation and Salient Object Detection

ResearchDGX agent

arXiv:2503.00747v2 Announce Type: replace Abstract: Light field cameras capture multi-view observations within a single exposure. However, existing studies are typically tailored to specific LF repres

LIDSA: Cognitive Arbitration for Signal-Free Autonomous Intersection Management

AgentsDGX agent

arXiv:2605.12321v2 Announce Type: replace Abstract: Large language models (LLMs) show strong potential for Intelligent Transportation Systems (ITS), particularly in tasks requiring situational reasoni

LightReasoner: Can Small Language Models Teach Large Language Models Reasoning?

ResearchDGX agent

arXiv:2510.07962v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated remarkable progress in reasoning, often through supervised fine-tuning (SFT). However, SFT is resourc

Linear Dynamics in the RLVR Training of Large Language Models

Model ReleasesDGX agent

arXiv:2601.04537v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven significant performance gains in reasoning-oriented large language models (LL

LLM Readiness Harness: Evaluation, Observability, and CI Gates for LLM/RAG Applications

Model ReleasesDGX agent

arXiv:2603.27355v2 Announce Type: replace-cross Abstract: We present a readiness harness for LLM and RAG applications that turns evaluation into a deployment decision workflow. The system combines aut

LongVT: Incentivizing 'Thinking with Long Videos' via Native Tool Calling

Model ReleasesDGX agent

arXiv:2511.20785v3 Announce Type: replace Abstract: Large multimodal models (LMMs) have shown great potential for video reasoning with textual Chain-of-Thought. However, they remain vulnerable to hall

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming

Local AiDGX agent

arXiv:2605.21652v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have significantly advanced medical visual question answering, yet their performance in ultrasound remains suboptimal. In

Lower Bounds for Advection-Diffusion Equations: An Exploration with AI-Generated Proofs

AgentsDGX agent

arXiv:2605.20623v1 Announce Type: cross Abstract: We establish explicit lower bounds for advection-diffusion equations in three settings: a polynomial ot H^{-1} bound for inviscid shears with uin L^in

LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model

Model ReleasesDGX agent

arXiv:2605.22089v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising framework for end-to-end autonomous driving. However, existing VLAs typically rely on sp

M3: Conversational LLMs Simplify Secure Clinical Data Access, Understanding, and Analysis

Model ReleasesDGX agent

arXiv:2507.01053v4 Announce Type: replace-cross Abstract: Large-scale clinical databases offer opportunities for medical research, but their complexity creates barriers to effective use. The Medical I

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles

Model ReleasesDGX agent

arXiv:2605.22177v1 Announce Type: cross Abstract: The proliferation of large language models (LLMs) and modular skills has endowed autonomous agents with increasingly powerful capabilities. Existing f

MagicFuse: Single Image Fusion for Visual and Semantic Reinforcement

TutorialsDGX agent

arXiv:2602.01760v2 Announce Type: replace Abstract: This paper focuses on a highly practical scenario: how to continue benefiting from the advantages of multi-modal image fusion under harsh conditions

Making the Discrete Continuous: Synthetic RAW Augmentations for Fine-Grained Evaluation of Person Detection Performance in Low Light

SafetyDGX agent

arXiv:2605.22455v1 Announce Type: new Abstract: Real-world deployment of AI vision models is both fueled and limited by the data available for training and testing. Real datasets are sparse and uneven

MAP4TS: A Multi-Aspect Prompting Framework for Time-Series Forecasting with Large Language Models

Model ReleasesDGX agent

arXiv:2510.23090v2 Announce Type: replace Abstract: Recent advances have investigated the use of pretrained large language models (LLMs) for time-series forecasting by aligning numerical inputs with L

Mapping Tomato Cropping Systems in California Using AlphaEarth Geospatial Embeddings and Deep Learning Analysis

SafetyDGX agent

arXiv:2605.21804v1 Announce Type: cross Abstract: Field-scale crop maps support supply-chain forecasting and policy, yet statewide crop identification still often depends on retrospective surveys or r

MARS: Modular Agent with Reflective Search for Automated AI Research

AgentsDGX agent

arXiv:2602.02660v3 Announce Type: replace Abstract: A critical bottleneck in automating AI research is the execution of complex machine learning engineering (MLE) tasks. MLE differs from general softw

MaSC: A Masked Similarity Metric for Evaluating Concept-Driven Generation

Model ReleasesDGX agent

arXiv:2605.22469v1 Announce Type: new Abstract: Evaluating single-concept personalization in text-to-image diffusion requires measuring both concept preservation, which captures identity fidelity to a

Matching with Deliberation: Test-Time Evolutionary Hierarchical Multi-Agents for Zero-Shot Compositional Image Retrieval

Model ReleasesDGX agent

arXiv:2605.22478v1 Announce Type: new Abstract: Zero-Shot Compositional Image Retrieval (ZS-CIR) requires both preserving the visual continuity of the reference image and faithfully executing the sema

MAVEN: A Multi-stage Agentic Annotation Pipeline for Video Reasoning Tasks

Model ReleasesDGX agent

arXiv:2605.21917v1 Announce Type: new Abstract: Training Vision Language Models (VLMs) for video event reasoning requires high-quality structured annotations capturing not only what happened, but when

Mind the Gaps: Multi-Robot Feedback-Driven Ergodic Coverage in Unknown Environments

ResearchDGX agent

arXiv:2605.21719v1 Announce Type: new Abstract: In this work, we address the problem of multi-robot adaptive coverage, where teams of robots perform dynamic sampling by continuously adjusting their po

MLLMs Know When Before Speaking: Revealing and Recovering Temporal Grounding via Attention Cues

Model ReleasesDGX agent

arXiv:2605.21954v1 Announce Type: new Abstract: Video temporal grounding (VTG), which localizes the start and end times of a queried event in an untrimmed video, is a key test of whether multimodal la

MM-Conv: A Multimodal Dataset and Benchmark for Context-Aware Grounding in 3D Dialogue

Model ReleasesDGX agent

arXiv:2605.21796v1 Announce Type: cross Abstract: Grounding language in the physical world requires AI systems to interpret references that emerge dynamically during conversation. While current vision

Modeling Emotional Dynamics in Agent-to-Agent Interactions on Moltbook

SafetyDGX agent

arXiv:2605.20442v1 Announce Type: cross Abstract: Generative AI systems are increasingly deployed as interactive agents in online environments, such as a social network called Moltbook. In Moltbook, l

Modeling Pathology-Like Behavioral Patterns in Language Models Through Behavioral Fine-Tuning

SafetyDGX agent

arXiv:2605.22356v1 Announce Type: new Abstract: Large language models are increasingly used as computational tools for modeling human-like behavior. We introduce a behavioral induction framework that

Moment-Reenacting: Inverse Motion Degradation with Cross-shutter Guidance

ApplicationsDGX agent

arXiv:2605.22423v1 Announce Type: new Abstract: Motion degradation, manifested as blur in global shutter (GS) images or rolling shutter (RS) distortion in RS counterparts, remains a fundamental challe

Moral Semantics Survive Machine Translation: Cross-Lingual Evidence from Moral Foundations Corpora

SafetyDGX agent

arXiv:2605.22660v1 Announce Type: new Abstract: Moral language is subtle and culturally variable, making it difficult to translate faithfully across languages. Idiomatic expressions, slang, and cultur

More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts

ResearchDGX agent

arXiv:2605.22641v1 Announce Type: new Abstract: Detecting Schwartz values in political text is difficult because implicit cues often depend on surrounding arguments and fine-grained distinctions betwe

MoSA: Motion-constrained Stress Adaptation for Mitigating Real-to-Sim Gap in Continuum Dynamics via Learning Residual Anisotropy

ApplicationsDGX agent

arXiv:2605.22597v1 Announce Type: cross Abstract: Learning real-world dynamics from visual observations is crucial for various domains. A common strategy is to calibrate simulators by estimating physi

MotiMotion: Motion-Controlled Video Generation with Visual Reasoning

Model ReleasesDGX agent

arXiv:2605.22818v1 Announce Type: new Abstract: Current motion-controlled image-to-video generation models rigidly follow user-provided trajectories that are often sparse, imprecise, and causally inco

Motion Design for Grasp-Based Dynamic Locomotion in Microgravity

ResearchDGX agent

arXiv:2605.21704v1 Announce Type: new Abstract: Locomotion in microgravity often relies on sparsely and irregularly arranged anchors, motivating grasp-based mobility with multiple limbs. In this setti

MotionDPS: Motion-Compensated 3D Brain MRI Reconstruction

ResearchDGX agent

arXiv:2605.22121v1 Announce Type: new Abstract: Magnetic resonance imaging (MRI) is highly susceptible to patient motion due to its relatively long acquisition times and the fact that data are acquire

MOTOR: A Multimodal Dataset for Two-Wheeler Rider Behavior Understanding

Model ReleasesDGX agent

arXiv:2605.22550v1 Announce Type: new Abstract: Two-wheelers account for a disproportionately high share of road fatalities in the Global South. Research on two-wheeler rider behavior, however, lags f

MRecover: A Conditional Generative Model for Recovering Motion-Corrupted MR images Using AI Generated Contrast

ResearchDGX agent

arXiv:2605.21669v1 Announce Type: new Abstract: Hippocampal subfield segmentation requires high-resolution T2w turbo spin echo (TSE) MRI, yet this sequence is susceptible to motion artifacts, leading

← Previous
1…620621622623624…1034
Next →