AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
Human
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
14 Apr 2026

FlowBind: Efficient Any-to-Any Generation with Bidirectional Flows

ResearchDGX agent

arXiv:2512.15420v2 Announce Type: replace Abstract: Any-to-any generation seeks to translate between arbitrary subsets of modalities, enabling flexible cross-modal synthesis. Despite recent success, e

FlowCoMotion: Text-to-Motion Generation via Token-Latent Flow Modeling

SafetyDGX agent

arXiv:2604.11083v1 Announce Type: cross Abstract: Text-to-motion generation is driven by learning motion representations for semantic alignment with language. Existing methods rely on either continuou

FlowHijack: A Dynamics-Aware Backdoor Attack on Flow-Matching Vision-Language-Action Models

ResearchDGX agent

arXiv:2604.09651v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are emerging as a cornerstone for robotics, with flow-matching policies like pi_0 showing great promise in generatin

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

FlowPalm: Optical Flow Driven Non-Rigid Deformation for Geometrically Diverse Palmprint Generation

Model ReleasesDGX agent

arXiv:2604.09989v1 Announce Type: cross Abstract: Recently, synthetic palmprints have been increasingly used as substitutes for real data to train recognition models. To be effective, such synthetic d

FM-Agent: Scaling Formal Methods to Large Systems via LLM-Based Hoare-Style Reasoning

AgentsDGX agent

arXiv:2604.11556v1 Announce Type: cross Abstract: LLM-assisted software development has become increasingly prevalent, and can generate large-scale systems, such as compilers. It becomes crucial to st

FM-SIREN & FM-FINER: Implicit Neural Representation Using Nyquist-based Orthogonality

ResearchDGX agent

arXiv:2509.23438v3 Announce Type: replace Abstract: Existing periodic activation-based implicit neural representation (INR) networks, such as SIREN and FINER, suffer from hidden feature redundancy, wh

ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling

ResearchDGX agent

arXiv:2603.22911v2 Announce Type: replace-cross Abstract: Due to the great saving of computation and memory overhead, token compression has become a research hot-spot for MLLMs and achieved remarkable

Fourier-KAN-Mamba: A Novel State-Space Equation Approach for Time-Series Anomaly Detection

ApplicationsDGX agent

arXiv:2511.15083v2 Announce Type: replace Abstract: Time-series anomaly detection plays a critical role in numerous real-world applications, including industrial monitoring and fault diagnosis. Recent

FPBench: A Comprehensive Benchmark of Multimodal Large Language Models for Fingerprint Analysis

Model ReleasesDGX agent

arXiv:2512.18073v2 Announce Type: replace Abstract: Multimodal LLMs (MLLMs) are capable of performing complex data analysis, visual question answering, generation, and reasoning tasks. However, their

FRAMER: Frequency-Aligned Self-Distillation with Adaptive Modulation Leveraging Diffusion Priors for Real-World Image Super-Resolution

SafetyDGX agent

arXiv:2512.01390v3 Announce Type: replace Abstract: Real-image super-resolution (Real-ISR) seeks to recover HR images from LR inputs with mixed, unknown degradations. While diffusion models surpass GA

FREE-Switch: Frequency-based Dynamic LoRA Switch for Style Transfer

SafetyDGX agent

arXiv:2604.10023v1 Announce Type: cross Abstract: With the growing availability of open-sourced adapters trained on the same diffusion backbone for diverse scenes and objects, combining these pretrain

FreeScale: Scaling 3D Scenes via Certainty-Aware Free-View Generation

ApplicationsDGX agent

arXiv:2604.10512v1 Announce Type: new Abstract: The development of generalizable Novel View Synthesis (NVS) models is critically limited by the scarcity of large-scale training data featuring diverse

From Agent Loops to Structured Graphs:A Scheduler-Theoretic Framework for LLM Agent Execution

Model ReleasesDGX agent

arXiv:2604.11378v1 Announce Type: new Abstract: The dominant paradigm for building LLM based agents is the Agent Loop, an iterative cycle where a single language model decides what to do next by readi

From Answers to Arguments: Toward Trustworthy Clinical Diagnostic Reasoning with Toulmin-Guided Curriculum Goal-Conditioned Learning

SafetyDGX agent

arXiv:2604.11137v1 Announce Type: new Abstract: The integration of Large Language Models (LLMs) into clinical decision support is critically obstructed by their opaque and often unreliable reasoning.

From Attribution to Action: A Human-Centered Application of Activation Steering

ResearchDGX agent

arXiv:2604.11467v1 Announce Type: new Abstract: Explainable AI (XAI) methods reveal which features influence model predictions, yet provide limited means for practitioners to act on these explanations

From Decision Trees to Boolean Logic: A Fast and Unified SHAP Algorithm

HardwareDGX agent

arXiv:2511.09376v2 Announce Type: replace Abstract: SHapley Additive exPlanations (SHAP) is a key tool for interpreting decision tree ensembles by assigning contribution values to features. It is wide

From GPT-3 to GPT-5: Mapping their capabilities, scope, limitations, and consequences

Model ReleasesDGX agent

arXiv:2604.10332v1 Announce Type: new Abstract: We present the progress of the GPT family from GPT-3 through GPT-3.5, GPT-4, GPT-4 Turbo, GPT-4o, GPT-4.1, and the GPT-5 family. Our work is comparative

From Helpful to Trustworthy: LLM Agents for Pair Programming

AgentsDGX agent

arXiv:2604.10300v1 Announce Type: cross Abstract: LLM-based coding agents are increasingly used to generate code, tests, and documentation. Still, their outputs can be plausible yet misaligned with de

From Perception to Autonomous Computational Modeling: A Multi-Agent Approach

Model ReleasesDGX agent

arXiv:2604.06788v2 Announce Type: replace-cross Abstract: We present a solver-agnostic framework in which coordinated large language model (LLM) agents autonomously execute the complete computational

From Perception to Planning: Evolving Ego-Centric Task-Oriented Spatiotemporal Reasoning via Curriculum Learning

ResearchDGX agent

arXiv:2604.10517v1 Announce Type: new Abstract: Modern vision-language models achieve strong performance in static perception, but remain limited in the complex spatiotemporal reasoning required for e

From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments

ResearchDGX agent

arXiv:2603.23964v2 Announce Type: replace Abstract: The remarkable progress of reinforcement learning (RL) is intrinsically tied to the environments used to train and evaluate artificial agents. Movin

From Query to Counsel: Structured Reasoning with a Multi-Agent Framework and Dataset for Legal Consultation

AgentsDGX agent

arXiv:2604.10470v1 Announce Type: cross Abstract: Legal consultation question answering (Legal CQA) presents unique challenges compared to traditional legal QA tasks, including the scarcity of high-qu

From Recency Bias to Stable Convergence Block Kaczmarz Methods for Online Preference Learning in Matchmaking Applications

SafetyDGX agent

arXiv:2604.09964v1 Announce Type: new Abstract: We present a family of Kaczmarz-based preference learning algorithms for real-time personalized matchmaking in reciprocal recommender systems. Post-step

From Redaction to Restoration: Deep Learning for Medical Image Anonymization and Reconstruction

ResearchDGX agent

arXiv:2604.11376v1 Announce Type: cross Abstract: Removing patient-specific information from medical images is crucial to enable sharing and open science without compromising patient identities. Howev

From Scalars to Tensors: Declared Losses Recover Epistemic Distinctions That Neutrosophic Scalars Cannot Express

Model ReleasesDGX agent

arXiv:2604.09602v1 Announce Type: new Abstract: Leyva-Vazquez and Smarandache (2025) demonstrated that neutrosophic T/I/F evaluation, where Truth, Indeterminacy, and Falsity are independent dimensions

From Speech-to-Spatial: Grounding Utterances on A Live Shared View with Augmented Reality

ResearchDGX agent

arXiv:2602.03059v2 Announce Type: replace-cross Abstract: We introduce Speech-to-Spatial, a referent disambiguation framework that converts verbal remote-assistance instructions into spatially grounde

From Theory to Protocol: Executable Frameworks for Creative Emergence and Strategic Foresight

Model ReleasesDGX agent

arXiv:2604.09597v1 Announce Type: cross Abstract: Creativity and strategic foresight have been extensively studied through descriptive theories -- Koestler's bisociation (1964), de Bono's lateral thin

From Topology to Trajectory: LLM-Driven World Models For Supply Chain Resilience

Model ReleasesDGX agent

arXiv:2604.11041v1 Announce Type: new Abstract: Semiconductor supply chains face unprecedented resilience challenges amidst global geopolitical turbulence. Conventional Large Language Model (LLM) plan

From Translation to Superset: Benchmark-Driven Evolution of a Production AI Agent from Rust to Python

Model ReleasesDGX agent

arXiv:2604.11518v1 Announce Type: cross Abstract: Cross-language migration of large software systems is a persistent engineering challenge, particularly when the source codebase evolves rapidly. We pr

From UAV Imagery to Agronomic Reasoning: A Multimodal LLM Benchmark for Plant Phenotyping

Model ReleasesDGX agent

arXiv:2604.09907v1 Announce Type: cross Abstract: To improve crop genetics, high-throughput, effective and comprehensive phenotyping is a critical prerequisite. While such tasks were traditionally per

From Understanding to Creation: A Prerequisite-Free AI Literacy Course with Technical Depth Across Majors

AgentsDGX agent

arXiv:2604.09634v1 Announce Type: cross Abstract: Most AI literacy courses for non-technical undergraduates emphasize conceptual breadth over technical depth. This paper describes UNIV 182, a prerequi

Frugal Knowledge Graph Construction with Local LLMs: A Zero-Shot Pipeline, Self-Consistency and Wisdom of Artificial Crowds

Model ReleasesDGX agent

arXiv:2604.11104v1 Announce Type: new Abstract: This paper presents an empirical study of a multi-model zero-shot pipeline for knowledge graph construction and exploitation, executed entirely through

FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language Models

Model ReleasesDGX agent

arXiv:2509.20624v5 Announce Type: replace-cross Abstract: Autoregressive language models (ARMs) deliver strong likelihoods, but are inherently serial: they generate one token per forward pass, which l

Fusion Complexity Inversion: Why Simpler Cross View Modules Outperform SSMs and Cross View Attention Transformers for Pasture Biomass Regression

Model ReleasesDGX agent

arXiv:2603.07819v3 Announce Type: replace Abstract: Accurate estimation of pasture biomass from agricultural imagery is critical for sustainable livestock management, yet existing methods are limited

Gait Recognition with Temporal Kolmogorov-Arnold Networks

Local AiDGX agent

arXiv:2604.09990v1 Announce Type: new Abstract: Gait recognition is a biometric modality that identifies individuals from their characteristic walking patterns. Unlike conventional biometric traits, g

GAMBIT: A Gamified Jailbreak Framework for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2601.03416v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have become widely deployed, yet their safety alignment remains fragile under adversarial inputs. Previous

GameplayQA: A Benchmarking Framework for Decision-Dense POV-Synced Multi-Video Understanding of 3D Virtual Agents

AgentsDGX agent

arXiv:2603.24329v2 Announce Type: replace-cross Abstract: Multimodal LLMs are increasingly deployed as perceptual backbones for autonomous agents in 3D environments, from robotics to virtual worlds. T

GaNI: Global and Near Field Illumination Aware Neural Inverse Rendering

ResearchDGX agent

arXiv:2403.15651v5 Announce Type: replace Abstract: In this paper, we present GaNI, a Global and Near-field Illumination-aware neural inverse rendering technique that can reconstruct geometry, albedo,

GanitLLM: Difficulty-Aware Bengali Mathematical Reasoning through Curriculum-GRPO

ResearchDGX agent

arXiv:2601.06767v2 Announce Type: replace-cross Abstract: We present a Bengali mathematical reasoning model called GanitLLM (named after the Bangla word for mathematics, 'Ganit'), together with a new

GazeVaLM: A Multi-Observer Eye-Tracking Benchmark for Evaluating Clinical Realism in AI-Generated X-Rays

Model ReleasesDGX agent

arXiv:2604.11653v1 Announce Type: new Abstract: We introduce GazeVaLM, a public eye-tracking dataset for studying clinical perception during chest radiograph authenticity assessment. The dataset compr

General-purpose LLMs as Models of Human Driver Behavior: The Case of Simplified Merging

Model ReleasesDGX agent

arXiv:2604.09609v1 Announce Type: new Abstract: Human behavior models are essential as behavior references and for simulating human agents in virtual safety assessment of automated vehicles (AVs), yet

General365: Benchmarking General Reasoning in Large Language Models Across Diverse and Challenging Tasks

Model ReleasesDGX agent

arXiv:2604.11778v1 Announce Type: cross Abstract: Contemporary large language models (LLMs) have demonstrated remarkable reasoning capabilities, particularly in specialized domains like mathematics an

Generalizable Deepfake Detection Based on Forgery-aware Layer Masking and Multi-artifact Subspace Decomposition

Model ReleasesDGX agent

arXiv:2601.01041v3 Announce Type: replace Abstract: Deepfake detection remains highly challenging, particularly in cross-dataset scenarios and complex real-world settings. This challenge mainly arises

Generating Hadamard matrices with transformers

ResearchDGX agent

arXiv:2604.11101v1 Announce Type: cross Abstract: We present a new method for constructing Hadamard matrices that combines transformer neural networks with local search in the PatternBoost framework.

Generating High Quality Synthetic Data for Dutch Medical Conversations

ResearchDGX agent

arXiv:2604.09645v1 Announce Type: cross Abstract: Medical conversations offer insights into clinical communication often absent from Electronic Health Records. However, developing reliable clinical Na

Generating Multiple-Choice Knowledge Questions with Interpretable Difficulty Estimation using Knowledge Graphs and Large Language Models

ApplicationsDGX agent

arXiv:2604.10748v1 Announce Type: cross Abstract: Generating multiple-choice questions (MCQs) with difficulty estimation remains challenging in automated MCQ-generation systems used in adaptive, AI-as

Generation-Augmented Generation: A Plug-and-Play Framework for Private Knowledge Injection in Large Language Models

Model ReleasesDGX agent

arXiv:2601.08209v3 Announce Type: replace Abstract: In domains such as materials science, biomedicine, and finance, high-stakes deployment of large language models (LLMs) requires injecting private, d

Generative Design for Direct-to-Chip Liquid Cooling for Data Centers

HardwareDGX agent

arXiv:2604.10941v1 Announce Type: cross Abstract: Rapid growth in artificial intelligence (AI) workloads is driving up data center power densities, increasing the need for advanced thermal management.

Generative Path-Finding Method for Wasserstein Gradient Flow

ResearchDGX agent

arXiv:2604.11519v1 Announce Type: new Abstract: Wasserstein gradient flows (WGFs) describe the evolution of probability distributions in Wasserstein space as steepest descent dynamics for a free energ

Generative UI: LLMs are Effective UI Generators

ResearchDGX agent

arXiv:2604.09577v1 Announce Type: cross Abstract: AI models excel at creating content, but typically render it with static, predefined interfaces. Specifically, the output of LLMs is often a markdown

GenProve: Learning to Generate Text with Fine-Grained Provenance

SafetyDGX agent

arXiv:2601.04932v2 Announce Type: replace Abstract: Large language models (LLM) often hallucinate, and while adding citations is a common solution, it is frequently insufficient for accountability as

GenTac: Generative Modeling and Forecasting of Soccer Tactics

Model ReleasesDGX agent

arXiv:2604.11786v1 Announce Type: new Abstract: Modeling open-play soccer tactics is a formidable challenge due to the stochastic, multi-agent nature of the game. Existing computational approaches typ

GeoArena: Evaluating Open-World Geographic Reasoning in Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2509.04334v4 Announce Type: replace Abstract: Geographic reasoning is a fundamental cognitive capability that requires models to infer plausible locations by synthesizing visual evidence with sp

GeoFormer: A Lightweight Swin Transformer for Joint Building Height and Footprint Estimation from Sentinel Imagery

Model ReleasesDGX agent

arXiv:2602.09932v2 Announce Type: replace Abstract: Building height (BH) and footprint (BF) are fundamental urban morphological parameters required by climate modelling, disaster-risk assessment, and

GeoMeld: Toward Semantically Grounded Foundation Models for Remote Sensing

SafetyDGX agent

arXiv:2604.10591v1 Announce Type: cross Abstract: Effective foundation modeling in remote sensing requires spatially aligned heterogeneous modalities coupled with semantically grounded supervision, ye

Geometry-Aware Localized Watermarking for Copyright Protection in Embedding-as-a-Service

Model ReleasesDGX agent

arXiv:2604.11344v1 Announce Type: cross Abstract: Embedding-as-a-Service (EaaS) has become an important semantic infrastructure for natural language and multimedia applications, but it is highly vulne

GeomPrompt: Geometric Prompt Learning for RGB-D Semantic Segmentation Under Missing and Degraded Depth

ResearchDGX agent

arXiv:2604.11585v1 Announce Type: new Abstract: Multimodal perception systems for robotics and embodied AI often assume reliable RGB-D sensing, but in practice, depth is frequently missing, noisy, or

Geoparsing: Diagram Parsing for Plane and Solid Geometry with a Unified Formal Language

ApplicationsDGX agent

arXiv:2604.11600v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but continue to struggle with geometric reasoning, primarily due to the perce

GIANTS: Generative Insight Anticipation from Scientific Literature

Model ReleasesDGX agent

arXiv:2604.09793v1 Announce Type: cross Abstract: Scientific breakthroughs often emerge from synthesizing prior ideas into novel contributions. While language models (LMs) show promise in scientific d

GIF: A Conditional Multimodal Generative Framework for IR Drop Imaging in Chip Layouts

TutorialsDGX agent

arXiv:2604.09999v1 Announce Type: new Abstract: IR drop analysis is essential in physical chip design to ensure the power integrity of on-chip power delivery networks. Traditional Electronic Design Au

← Previous
1…944945946947948…989
Next →