AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
Human
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
10 Apr 2026

Thinking in Graphs with CoMAP: A Shared Visual Workspace for Designing Project-Based Learning

ResearchDGX agent

arXiv:2604.06200v1 Announce Type: cross Abstract: Designing project-based learning (PBL) demands managing highly interdependent components, a task that both traditional linear tools and purely convers

Thompson Sampling for Infinite-Horizon Discounted Decision Processes

Model ReleasesDGX agent

arXiv:2405.08253v3 Announce Type: replace-cross Abstract: This paper develops a viable notion of learning for sampling-based algorithms that applies in broader settings than previously considered. Mor

Through the Magnifying Glass: Adaptive Perception Magnification for Hallucination-Free VLM Decoding

ResearchDGX agent

arXiv:2503.10183v4 Announce Type: replace Abstract: Existing vision-language models (VLMs) often suffer from visual hallucination, where the generated responses contain inaccuracies that are not groun

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Tight Convergence Rates for Online Distributed Linear Estimation with Adversarial Measurements

Model ReleasesDGX agent

arXiv:2604.06282v1 Announce Type: cross Abstract: We study mean estimation of a random vector X in a distributed parameter-server-worker setup. Worker i observes samples of a_i^op X, where $a_

Time-Series Classification with Multivariate Statistical Dependence Features

ResearchDGX agent

arXiv:2604.06537v1 Announce Type: new Abstract: In this paper, we propose a novel framework for non-stationary time-series analysis that replaces conventional correlation-based statistics with direct

Tool-MCoT: Tool Augmented Multimodal Chain-of-Thought for Content Safety Moderation

SafetyDGX agent

arXiv:2604.06205v1 Announce Type: cross Abstract: The growth of online platforms and user content requires strong content moderation systems that can handle complex inputs from various media types. Wh

Tool Retrieval Bridge: Aligning Vague Instructions with Retriever Preferences via Bridge Model

Model ReleasesDGX agent

arXiv:2604.07816v1 Announce Type: new Abstract: Tool learning has emerged as a promising paradigm for large language models (LLMs) to address real-world challenges. Due to the extensive and irregularl

TOOLCAD: Exploring Tool-Using Large Language Models in Text-to-CAD Generation with Reinforcement Learning

AgentsDGX agent

arXiv:2604.07960v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) is an expert-level task that relies on long-horizon reasoning and coherent modeling actions. Large Language Models (LLMs)

Toward a Tractability Frontier for Exact Relevance Certification

ResearchDGX agent

arXiv:2604.07349v1 Announce Type: cross Abstract: Exact relevance certification asks which coordinates are necessary to determine the optimal action in a coordinate-structured decision problem. The tr

Toward a universal foundation model for graph-structured data

Model ReleasesDGX agent

arXiv:2604.06391v1 Announce Type: cross Abstract: Graphs are a central representation in biomedical research, capturing molecular interaction networks, gene regulatory circuits, cell--cell communicati

Toward Memory-Aided World Models: Benchmarking via Spatial Consistency

Model ReleasesDGX agent

arXiv:2505.22976v2 Announce Type: replace-cross Abstract: The ability to simulate the world in a spatially consistent manner is a crucial requirements for effective world models. Such a model enables

Toward Personalized Darts Training: A Data-Driven Framework Based on Skeleton-Based Biomechanical Analysis and Motion Modeling

Local AiDGX agent

arXiv:2604.01130v3 Announce Type: replace Abstract: As sports training becomes more data-driven, traditional dart coaching based mainly on experience and visual observation is increasingly inadequate

Toward Reducing Unproductive Container Moves: Predicting Service Requirements and Dwell Times

ResearchDGX agent

arXiv:2604.06251v1 Announce Type: new Abstract: This article presents the results of a data science study conducted at a container terminal, aimed at reducing unproductive container moves through the

Towards Accurate and Calibrated Classification: Regularizing Cross-Entropy From A Generative Perspective

Model ReleasesDGX agent

arXiv:2604.06689v1 Announce Type: new Abstract: Accurate classification requires not only high predictive accuracy but also well-calibrated confidence estimates. Yet, modern deep neural networks (DNNs

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval

Model ReleasesDGX agent

arXiv:2512.08410v2 Announce Type: replace Abstract: Due to excessive memory overhead, most Multimodal Large Language Models (MLLMs) can only process videos of limited frames. In this paper, we propose

Towards Hierarchical Multi-Step Reward Models for Enhanced Reasoning in Large Language Models

ResearchDGX agent

arXiv:2503.13551v5 Announce Type: replace Abstract: Recent studies show that Large Language Models (LLMs) achieve strong reasoning capabilities through supervised fine-tuning or reinforcement learning

Towards Privacy-Preserving Large Language Model: Text-free Inference Through Alignment and Adaptation

SafetyDGX agent

arXiv:2604.06831v1 Announce Type: cross Abstract: Current LLM-based services typically require users to submit raw text regardless of its sensitivity. While intuitive, such practice introduces substan

Towards provable probabilistic safety for scalable embodied AI systems

SafetyDGX agent

arXiv:2506.05171v3 Announce Type: replace-cross Abstract: Embodied AI systems, comprising AI models and physical plants, are increasingly prevalent across various applications. Due to the rarity of sy

Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous Behavior Traces

Model ReleasesDGX agent

arXiv:2604.08362v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) has illuminated the potential for a general-purpose user simulator. However, existing benchmarks remain co

Towards Resilient Intrusion Detection in CubeSats: Challenges, TinyML Solutions, and Future Directions

AgentsDGX agent

arXiv:2604.06411v1 Announce Type: cross Abstract: CubeSats have revolutionized access to space by providing affordable and accessible platforms for research and education. However, their reliance on C

Towards Robust Content Watermarking Against Removal and Forgery Attacks

ResearchDGX agent

arXiv:2604.06662v1 Announce Type: cross Abstract: Generated contents have raised serious concerns about copyright protection, image provenance, and credit attribution. A potential solution for these p

Towards the Development of an LLM-Based Methodology for Automated Security Profiling in Compliance with Ukrainian Cybersecurity Regulations

SafetyDGX agent

arXiv:2604.06274v1 Announce Type: cross Abstract: In recent years, the pace of development of information technology in various areas has increased drastically, forcing cybersecurity specialists to co

ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway

Model ReleasesDGX agent

arXiv:2604.06264v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled molecular reasoning for property prediction. However, toxicity arises from complex biolog

TR-EduVSum: A Turkish-Focused Dataset and Consensus Framework for Educational Video Summarization

Model ReleasesDGX agent

arXiv:2604.07553v1 Announce Type: new Abstract: This study presents a framework for generating the gold-standard summary fully automatically and reproducibly based on multiple human summaries of Turki

TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories

Model ReleasesDGX agent

arXiv:2604.07223v1 Announce Type: cross Abstract: As large language models (LLMs) evolve from static chatbots into autonomous agents, the primary vulnerability surface shifts from final outputs to int

Tracking Adaptation Time: Metrics for Temporal Distribution Shift

TutorialsDGX agent

arXiv:2604.07266v1 Announce Type: new Abstract: Evaluating robustness under temporal distribution shift remains an open challenge. Existing metrics quantify the average decline in performance, but fai

Training Data Size Sensitivity in Unsupervised Rhyme Recognition

Model ReleasesDGX agent

arXiv:2604.08156v1 Announce Type: new Abstract: Rhyme is deceptively intuitive: what is or is not a rhyme is constructed historically, scholars struggle with rhyme classification, and people disagree

Training-free Spatially Grounded Geometric Shape Encoding (Technical Report)

ResearchDGX agent

arXiv:2604.07522v1 Announce Type: new Abstract: Positional encoding has become the de facto standard for grounding deep neural networks on discrete point-wise positions, and it has achieved remarkable

Transformer See, Transformer Do: Copying as an Intermediate Step in Learning Analogical Reasoning

ResearchDGX agent

arXiv:2604.06501v1 Announce Type: new Abstract: Analogical reasoning is a hallmark of human intelligence, enabling us to solve new problems by transferring knowledge from one situation to another. Yet

Transforming the Voice of the Customer: Large Language Models for Identifying Customer Needs

TutorialsDGX agent

arXiv:2503.01870v2 Announce Type: replace Abstract: Identifying customer needs (CNs) is fundamental to product innovation and marketing strategy. Yet for over thirty years, Voice-of-the-Customer (VOC)

TREASURE: The Visa Payment Foundation Model for High-Volume Transaction Understanding

ApplicationsDGX agent

arXiv:2511.19693v3 Announce Type: replace-cross Abstract: Payment networks form the backbone of modern commerce, generating high volumes of transaction records from daily activities. Properly modeling

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation

Model ReleasesDGX agent

arXiv:2604.07894v1 Announce Type: new Abstract: Personalized large language models (PLLMs) have garnered significant attention for their ability to align outputs with individual's needs and preference

TurboAgent: An LLM-Driven Autonomous Multi-Agent Framework for Turbomachinery Aerodynamic Design

AgentsDGX agent

arXiv:2604.06747v2 Announce Type: new Abstract: The aerodynamic design of turbomachinery is a complex and tightly coupled multi-stage process involving geometry generation, performance prediction, opt

TwinLoop: Simulation-in-the-Loop Digital Twins for Online Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2604.06610v1 Announce Type: cross Abstract: Decentralised online learning enables runtime adaptation in cyber-physical multi-agent systems, but when operating conditions change, learned policies

U-CECE: A Universal Multi-Resolution Framework for Conceptual Counterfactual Explanations

ResearchDGX agent

arXiv:2604.08295v1 Announce Type: cross Abstract: As AI models grow more complex, explainability is essential for building trust, yet concept-based counterfactual methods still face a trade-off betwee

UI-AGILE: Advancing GUI Agents with Effective Reinforcement Learning and Precise Inference-Time Grounding

AgentsDGX agent

arXiv:2507.22025v4 Announce Type: replace Abstract: The emergence of Multimodal Large Language Models (MLLMs) has driven significant advances in Graphical User Interface (GUI) agent capabilities. Neve

Uncertainty Estimation for Deep Reconstruction in Actuatic Disaster Scenarios with Autonomous Vehicles

AgentsDGX agent

arXiv:2604.06387v1 Announce Type: cross Abstract: Accurate reconstruction of environmental scalar fields from sparse onboard observations is essential for autonomous vehicles engaged in aquatic monito

Understanding Structured Financial Data with LLMs: A Case Study on Fraud Detection

ApplicationsDGX agent

arXiv:2512.13040v2 Announce Type: replace-cross Abstract: Detecting fraud in financial transactions typically relies on tabular models that demand heavy feature engineering to handle high-dimensional

Understanding Task Transfer in Vision-Language Models

TutorialsDGX agent

arXiv:2511.18787v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) perform well on multimodal benchmarks but lag behind humans and specialized models on visual perception tasks like dep

Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator

ResearchDGX agent

arXiv:2604.08121v1 Announce Type: new Abstract: Unified multimodal models integrating visual understanding and generation face a fundamental challenge: visual generation incurs substantially higher co

Unifying Speech Editing Detection and Content Localization via Prior-Enhanced Audio LLMs

Model ReleasesDGX agent

arXiv:2601.21463v2 Announce Type: replace-cross Abstract: Existing speech editing detection (SED) datasets are predominantly constructed using manual splicing or limited editing operations, resulting

UniLACT: Depth-Aware RGB Latent Action Learning for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2602.20231v2 Announce Type: replace-cross Abstract: Latent action representations learned from unlabeled videos have recently emerged as a promising paradigm for pretraining vision-language-acti

UniversalVTG: A Universal and Lightweight Foundation Model for Video Temporal Grounding

Model ReleasesDGX agent

arXiv:2604.08522v1 Announce Type: new Abstract: Video temporal grounding (VTG) is typically tackled with dataset-specific models that transfer poorly across domains and query styles. Recent efforts to

Unsupervised Neural Network for Automated Classification of Surgical Urgency Levels in Medical Transcriptions

ApplicationsDGX agent

arXiv:2604.06214v1 Announce Type: cross Abstract: Efficient classification of surgical procedures by urgency is paramount to optimize patient care and resource allocation within healthcare systems. Th

URMF: Uncertainty-aware Robust Multimodal Fusion for Multimodal Sarcasm Detection

SafetyDGX agent

arXiv:2604.06728v1 Announce Type: cross Abstract: Multimodal sarcasm detection (MSD) aims to identify sarcastic intent from semantic incongruity between text and image. Although recent methods have im

Validated Intent Compilation for Constrained Routing in LEO Mega-Constellations

Model ReleasesDGX agent

arXiv:2604.07264v1 Announce Type: cross Abstract: Operating LEO mega-constellations requires translating high-level operator intents ('reroute financial traffic away from polar links under 80 ms') int

VAREX: A Benchmark for Multi-Modal Structured Extraction from Documents

Model ReleasesDGX agent

arXiv:2603.15118v2 Announce Type: replace Abstract: We introduce VAREX (VARied-schema EXtraction), a benchmark for evaluating multimodal foundation models on structured data extraction from government

Variational Feature Compression for Model-Specific Representations

Model ReleasesDGX agent

arXiv:2604.06644v1 Announce Type: cross Abstract: As deep learning inference is increasingly deployed in shared and cloud-based settings, a growing concern is input repurposing, in which data submitte

VenusBench-Mobile: A Challenging and User-Centric Benchmark for Mobile GUI Agents with Capability Diagnostics

Model ReleasesDGX agent

arXiv:2604.06182v1 Announce Type: cross Abstract: Existing online benchmarks for mobile GUI agents remain largely app-centric and task-homogeneous, failing to reflect the diversity and instability of

Verify Before You Commit: Towards Faithful Reasoning in LLM Agents via Self-Auditing

Model ReleasesDGX agent

arXiv:2604.08401v1 Announce Type: cross Abstract: In large language model (LLM) agents, reasoning trajectories are treated as reliable internal beliefs for guiding actions and updating memory. However

VertAX: a differentiable vertex model for learning epithelial tissue mechanics

Model ReleasesDGX agent

arXiv:2604.06896v1 Announce Type: new Abstract: Epithelial tissues dynamically reshape through local mechanical interactions among cells, a process well captured by vertex models. Yet their many tunab

Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs

Model ReleasesDGX agent

arXiv:2509.08016v2 Announce Type: replace Abstract: Video Large Language Models (VideoLLMs) face a critical bottleneck: increasing the number of input frames to capture fine-grained temporal detail le

VisCoder2: Building Multi-Language Visualization Coding Agents

Model ReleasesDGX agent

arXiv:2510.23642v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently enabled coding agents capable of generating, executing, and revising visualization code. However, e

Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment

ApplicationsDGX agent

arXiv:2604.08212v1 Announce Type: new Abstract: General-purpose vision-language models demonstrate strong performance in everyday domains but struggle with specialized technical fields requiring preci

Vision-Language Navigation for Aerial Robots: Towards the Era of Large Language Models

Model ReleasesDGX agent

arXiv:2604.07705v1 Announce Type: new Abstract: Aerial vision-and-language navigation (Aerial VLN) aims to enable unmanned aerial vehicles (UAVs) to interpret natural language instructions and autonom

VisionClaw: Always-On AI Agents through Smart Glasses

AgentsDGX agent

arXiv:2604.03486v2 Announce Type: replace-cross Abstract: We present VisionClaw, an always-on wearable AI agent that integrates live egocentric perception with agentic task execution. Running on Meta

Visual prompting reimagined: The power of the Activation Prompts

Model ReleasesDGX agent

arXiv:2604.06440v1 Announce Type: cross Abstract: Visual prompting (VP) has emerged as a popular method to repurpose pretrained vision models for adaptation to downstream tasks. Unlike conventional mo

Visually-grounded Humanoid Agents

Model ReleasesDGX agent

arXiv:2604.08509v1 Announce Type: new Abstract: Digital human generation has been studied for decades and supports a wide range of real-world applications. However, most existing systems are passively

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning

SafetyDGX agent

arXiv:2604.08168v1 Announce Type: new Abstract: Vision-language-action (VLA) models have advanced robot manipulation through large-scale pretraining, but real-world deployment remains challenging due

VLMShield: Efficient and Robust Defense of Vision-Language Models against Malicious Prompts

SafetyDGX agent

arXiv:2604.06502v1 Announce Type: new Abstract: Vision-Language Models (VLMs) face significant safety vulnerabilities from malicious prompt attacks due to weakened alignment during visual integration.

← Previous
1…995996997998
Next →