AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
Research

From Baseline to Follow-Up: Counterfactual Spine DXA Image Synthesis in UK Biobank Using a Causal Hierarchical Variational Autoencoder

DGX agent

arXiv:2605.22649v1 Announce Type: new Abstract: Dual-energy X-ray absorptiometry (DXA) is widely used for large-scale skeletal assessment, yet learning controllable and interpretable factor-specific a

researcharxiv-cs-cv
22 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

From Correlation to Cause: A Five-Stage Methodology for Feature Analysis in Transformer Language Models

DGX agent

arXiv:2605.22462v1 Announce Type: new Abstract: We propose a five-stage methodology for causal feature analysis in transformer language models (probe design, feature extraction, causal validation, rob

researcharxiv-cs-cl
22 May 2026
Model Releases

From Parameters to Data: A Task-Parameter-Guided Fine-Tuning Pipeline for Efficient LLM Alignment

DGX agent

arXiv:2605.21558v1 Announce Type: cross Abstract: Adapting Large Language Models (LLMs) to specialized domains typically incurs high data and computational overhead. While prior efficiency efforts hav

model-releasesarxiv-cs-cl
22 May 2026
Research

From Reasoning Chains to Verifiable Subproblems: Curriculum Reinforcement Learning Enables Credit Assignment for LLM Reasoning

DGX agent

arXiv:2605.22074v1 Announce Type: cross Abstract: Reinforcement learning from verifiable rewards (RLVR) has shown strong promise for LLM reasoning, but outcome-based RLVR remains inefficient on hard p

researcharxiv-cs-cl
22 May 2026
Model Releases

From Recognition to Reasoning: Benchmarking and Enhancing MLLMs on Real-World Receipt Document Understanding

DGX agent

arXiv:2605.22413v1 Announce Type: new Abstract: Extracting structured information from visual documents (Visual Information Extraction, VIE) is a cornerstone of business automation. While recent Multi

model-releasesarxiv-cs-cv
22 May 2026
Research

From TF-IDF to Transformers: A Comparative and Ensemble Approach to Sentiment Classification

DGX agent

arXiv:2605.22003v1 Announce Type: new Abstract: Sentiment analysis, also referred to as opinion mining, primarily tries to extract opinion from any text-based data. In the context of movie reviews and

researcharxiv-cs-cl
22 May 2026
Agents

GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation

DGX agent

arXiv:2605.22036v1 Announce Type: new Abstract: Despite significant progress in Vision-Language Navigation (VLN), existing approaches still rely on dense RGB videos that produce excessive patch tokens

agentsarxiv-cs-cv
22 May 2026
Local Ai

GALAR-TemporalNet v2: Anatomy-Guided Dual-Branch Temporal Classification with Bidirectional Mamba and Dual-Graph GCN for Video Capsule Endoscopy -- after competition results

DGX agent

arXiv:2605.22209v1 Announce Type: new Abstract: Video Capsule Endoscopy (VCE) poses a challenging multi-label temporal classification problem, requiring simultaneous localization of 8 anatomical regio

local-aiarxiv-cs-cv
22 May 2026
Research

GazePrior: Zero-Shot AR/VR Eye Tracking via Learned 3D Gaze Reconstruction

DGX agent

arXiv:2605.22359v1 Announce Type: new Abstract: Eye tracking (ET) is a foundational technology for advanced AR/VR applications. However, training ET models for every new ET device is challenging: real

researcharxiv-cs-cv
22 May 2026
Agents

General Agentic Planning Through Simulative Reasoning with World Models

DGX agent

arXiv:2507.23773v3 Announce Type: replace-cross Abstract: What does it mean to plan? Current agentic systems, whether scaffolded workflows or end-to-end policies, rely on reactive decision-making: sel

agentsarxiv-cs-cl
22 May 2026
Safety

GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation

DGX agent

arXiv:2605.21605v1 Announce Type: new Abstract: Open-ended image generation is no longer a simple prompt-to-image problem. High-quality generation often requires an agent to combine a model's internal

safetyarxiv-cs-cv
22 May 2026
Applications

GenHAR: Generalizing Cross-domain Human Activity Recognition for Last-mile Delivery

DGX agent

arXiv:2605.22086v1 Announce Type: new Abstract: Human Activity Recognition (HAR) has shown remarkable effectiveness in various applications, such as smart healthcare and intelligent manufacturing. How

applicationsarxiv-cs-cv
22 May 2026
Research

Geometry-Adaptive Explainer for Faithful Dictionary-Based Interpretability under Distribution Shift

DGX agent

arXiv:2605.21849v1 Announce Type: cross Abstract: Mechanistic interpretability aims to explain a model's behavior by identifying causally responsible internal structures. Dictionary-based explainers s

researcharxiv-cs-cl
22 May 2026
Model Releases

GeoWeaver: Grounding Visual Tokens with Geometric Evidence before Scene Reasoning

DGX agent

arXiv:2605.22558v1 Announce Type: new Abstract: Spatio-temporal reasoning in vision-language models requires visual representations that preserve physical geometry rather than merely semantic appearan

model-releasesarxiv-cs-cv
22 May 2026
Applications

GesVLA: Gesture-Aware Vision-Language-Action Model Embedded Representations

DGX agent

arXiv:2605.22812v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general-purpose robot manipulation by unifying perception and action. However, exi

applicationsarxiv-cs-cv
22 May 2026
Model Releases

GHI: Graphormer over Conditioned Hypergraph Incidence for Aspect-Based Sentiment Analysis

DGX agent

arXiv:2605.22228v1 Announce Type: new Abstract: Aspect-based sentiment analysis (ABSA) requires models to bind sentiment evidence to the correct aspect, making it a natural testbed for fine-grained st

model-releasesarxiv-cs-cl
22 May 2026
Safety

GLeVE: Graph-Guided Lesion Grounding with Proposal Verification in 3D CT

DGX agent

arXiv:2605.22619v1 Announce Type: new Abstract: Grounding radiology report descriptions to 3D CT volumes is essential for verifiable clinical interpretation, yet remains challenging due to the semanti

safetyarxiv-cs-cv
22 May 2026
Safety

Governance by Construction for Generalist Agents

DGX agent

arXiv:2605.20874v1 Announce Type: new Abstract: Enterprise agents are increasingly expected to operate autonomously across tools and interfaces, yet production deployments require governance by constr

safetyarxiv-cs-ai
22 May 2026
Safety

Governance by Design: Architecting Agentic AI for Organizational Learning and Scalable Autonomy

DGX agent

arXiv:2605.20210v1 Announce Type: cross Abstract: Agentic AI systems - systems that can pursue goals through multi-step planning and tool-mediated action with limited direct supervision - are moving f

safetyarxiv-cs-ai
22 May 2026
Model Releases

GrandGuard: Taxonomy, Benchmark, and Safeguards for Elderly-Chatbot Interaction Safety

DGX agent

arXiv:2605.20203v1 Announce Type: cross Abstract: As older adults increasingly use LLM-based chatbots for companionship and assistance, a safety gap is emerging. Older adults may face vulnerabilities

model-releasesarxiv-cs-ai
22 May 2026
Research

Guided Trajectory Optimization with Sparse Scaling for Test-Time Diffusion

DGX agent

arXiv:2605.21907v1 Announce Type: new Abstract: The efficient Test-Time Scaling (TTS) paradigm offers a promising perspective for enhancing the generation performance of diffusion models. However, cur

researcharxiv-cs-cv
22 May 2026
Model Releases

H-Flow: Self-supervised Human Scene Flow via Physics-inspired Joint Multi-modal Learning

DGX agent

arXiv:2605.22629v1 Announce Type: new Abstract: Parametric human models capture global pose but cannot represent the non-rigid surface dynamics of clothing and soft tissue. Generic scene flow estimate

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Hallucination as Commitment Failure: Larger LLMs Misfire Despite Knowing the Answer

DGX agent

arXiv:2605.22007v1 Announce Type: new Abstract: Hallucination is often viewed as a direct consequence of missing knowledge: a model answers incorrectly when the correct answer is absent from its gener

model-releasesarxiv-cs-cl
22 May 2026
Safety

Harder to Defend: Towards Chinese Toxicity Attacks via Implicit Enhancement and Obfuscation Rewriting

DGX agent

arXiv:2605.22258v1 Announce Type: new Abstract: Large language models (LLMs) require robust toxicity evaluation beyond explicit wording. This setting remains underexplored in Chinese, where toxicity m

safetyarxiv-cs-cl
22 May 2026
Model Releases

HealthCraft: A Reinforcement Learning Safety Environment for Emergency Medicine

DGX agent

arXiv:2605.21496v1 Announce Type: cross Abstract: Frontier language models are being deployed into clinical workflows faster than the infrastructure to evaluate them safely. Static medical-QA benchmar

model-releasesarxiv-cs-cl
22 May 2026
Local Ai

Heartbeat-Bound Hierarchical Credentials: Cryptographic Revocation for AI Agent Swarms

DGX agent

arXiv:2605.20704v1 Announce Type: cross Abstract: Autonomous AI agents that spawn sub-agent swarms create a safety gap: existing credential revocation mechanisms, OAuth~2.0 introspection, OCSP, and W3

local-aiarxiv-cs-ai
22 May 2026
Safety

Hierarchical Variational Policies for Reward-Guided Diffusion

DGX agent

arXiv:2605.21661v1 Announce Type: cross Abstract: Adapting pretrained diffusion models to downstream objectives such as inverse problems often requires expensive test-time guidance or optimization. We

safetyarxiv-cs-cv
22 May 2026
Research

High Quality Embeddings for Horn Logic Reasoning

DGX agent

arXiv:2605.20467v1 Announce Type: new Abstract: Neural networks can be trained to rank the choices made by logical reasoners, resulting in more efficient searches for answers. A key step in this proce

researcharxiv-cs-ai
22 May 2026
Research

Higher Order Reasoning for Collaborative Communicationless Mobile Robot Operations

DGX agent

arXiv:2605.21901v1 Announce Type: new Abstract: In communicationless environments, multi-robot systems must operate without the constant information exchange that many coordination strategies typicall

researcharxiv-cs-ro
22 May 2026
Safety

How can reasoning capability empower the AI copilot robot in endoscopic surgery

DGX agent

arXiv:2605.22322v1 Announce Type: new Abstract: Reasoning capability has significantly advanced complex logical inference and robotic decision-making in general domains. However, its potential in the

safetyarxiv-cs-ro
22 May 2026
Tutorials

How to Build Marcus's Algebraic Mind: Algebro-Deterministic Substrate over Galois Fields

DGX agent

arXiv:2605.21379v2 Announce Type: cross Abstract: In The Algebraic Mind, Gary Marcus identified three components essential for any adequate cognitive architecture: operations over variables, recursive

tutorialsarxiv-cs-ai
22 May 2026
Model Releases

How Well Do Models Follow Visual Instructions? VIBE: A Systematic Benchmark for Visual Instruction-Driven Image Editing

DGX agent

arXiv:2602.01851v2 Announce Type: replace Abstract: Recent generative models have achieved remarkable progress in image editing. However, existing systems and benchmarks remain largely text-guided. In

model-releasesarxiv-cs-cv
22 May 2026
Tutorials

HUSKY: Humanoid Skateboarding System via Physics-Aware Whole-Body Control

DGX agent

arXiv:2602.03205v2 Announce Type: replace Abstract: While current humanoid whole-body control frameworks predominantly rely on the static environment assumptions, addressing tasks characterized by hig

tutorialsarxiv-cs-ro
22 May 2026
Model Releases

Hy-MT2: A Family of Fast, Efficient and Powerful Multilingual Translation Models in the Wild

DGX agent

arXiv:2605.22064v1 Announce Type: new Abstract: Hy-MT2 is a family of fast-thinking multilingual translation models designed for complex real-world scenarios. It includes three model sizes: 1.8B, 7B,

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

HyLoVQA: Dynamic Hypernetwork-Generated Low-Rank Adaptation for Continual Visual Question Answering

DGX agent

arXiv:2605.22035v1 Announce Type: cross Abstract: Continual Visual Question Answering (VQA) requires learning from non-stationary streams of visual inputs and questions while preserving past knowledge

model-releasesarxiv-cs-cl
22 May 2026
Applications

HyperBench: Standardizing and Scaling Synthetic Evaluation for Hyperspectral Super-Resolution

DGX agent

arXiv:2605.21671v1 Announce Type: cross Abstract: Hyperspectral super-resolution (HSR) reconstructs a high-spatial-resolution hyperspectral image by fusing a low-resolution hyperspectral image (LR-HSI

applicationsarxiv-cs-cv
22 May 2026
Local Ai

Hypergraph as Language

DGX agent

arXiv:2605.21858v1 Announce Type: new Abstract: Large language models (LLMs) have recently shown strong potential in modeling relational structures. However, existing approaches remain fundamentally g

local-aiarxiv-cs-cl
22 May 2026
Model Releases

IdioLink: Retrieving Meaning Beyond Words Across Idiomatic and Literal Expressions

DGX agent

arXiv:2605.22247v1 Announce Type: new Abstract: Idioms pose a fundamental challenge for language models, as their meaning cannot be inferred from surface form alone. Understanding such expressions, th

model-releasesarxiv-cs-cl
22 May 2026
Research

Imagine2Real: Towards Zero-shot Humanoid-Object Interaction via Video Generative Priors

DGX agent

arXiv:2605.22272v1 Announce Type: cross Abstract: Whole-body Humanoid-Object Interaction (HOI) is bottlenecked by the scarcity of high-fidelity 3D data. While video generative priors offer a promising

researcharxiv-cs-cv
22 May 2026
Applications

Impact of Atmospheric Turbulence and Pointing Error on Earth Observation

DGX agent

arXiv:2605.22268v1 Announce Type: cross Abstract: Earth Observation (EO) imagery is often degraded by atmospheric turbulence and pointing jitter; yet, these effects are rarely considered in datasets u

applicationsarxiv-cs-cv
22 May 2026
Research

Improved DDIM Sampling with Moment Matching Gaussian Mixtures

DGX agent

arXiv:2311.04938v5 Announce Type: replace Abstract: We propose using a Gaussian Mixture Model (GMM) as reverse transition operator (kernel) within the Denoising Diffusion Implicit Models (DDIM) framew

researcharxiv-cs-cv
22 May 2026
Agents

ImProver: Agent-Based Automated Proof Optimization

DGX agent

arXiv:2410.04753v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been used to generate formal proofs of mathematical theorems in proofs assistants such as Lean. However, we

agentsarxiv-cs-cl
22 May 2026
Research

Improving 3D Labeling in Self-Driving by Inferring Vehicle Information using Vision Language Models

DGX agent

arXiv:2605.21747v1 Announce Type: new Abstract: We present an approach to improve 3D vehicle labeling in self-driving applications through zero-shot inference of vehicle information, leveraging Vehicl

researcharxiv-cs-cv
22 May 2026
Research

Improving Viewpoint-Invariance and Temporal Consistency for Action Detection

DGX agent

arXiv:2605.22695v1 Announce Type: new Abstract: Viewpoint change invariance and action temporal consistency are critical aspects for the effective deployment of human action detection of untrimmed vid

researcharxiv-cs-cv
22 May 2026
Research

In Silico Modeling of the RAMPHO Buffer: Dissociating Informational and Energetic Masking via Phonetic Entropy in Deep Neural Networks

DGX agent

arXiv:2605.22465v1 Announce Type: new Abstract: The fundamental challenge of listening in multi-talker environments is a cognitive bottleneck, defined by the Ease of Language Understanding (ELU) model

researcharxiv-cs-cl
22 May 2026
Research

Industrial Dual-Arm Box Handling via Online Inertial Estimation and Convex Wrench Optimization

DGX agent

arXiv:2605.22021v1 Announce Type: new Abstract: Industrial robotic object handling often involves boxes and packages whose mass and center of mass are not known in advance. These uncertainties affect

researcharxiv-cs-ro
22 May 2026
Model Releases

InfVSR: Breaking Length Limits of Generic Video Super-Resolution

DGX agent

arXiv:2510.00948v2 Announce Type: replace Abstract: Real-world videos often extend over thousands of frames. Existing generative video super-resolution (VSR) approaches, however, face two persistent c

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

InnerQ: Hardware-Aware Tuning-Free Quantization of KV Cache for Large Language Models

DGX agent

arXiv:2602.23200v2 Announce Type: replace-cross Abstract: When transformer-based language models are deployed for text generation, most of the inference time is spent in the decoding stage, where outp

model-releasesarxiv-cs-cl
22 May 2026
← Previous
1…793794795796797…1311
Next →