AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,151 results
Agents

Scaling Beyond Context: A Survey of Multimodal Retrieval-Augmented Generation for Document Understanding

DGX agent

arXiv:2510.15253v3 Announce Type: replace Abstract: Document understanding is critical for applications from financial analysis to scientific discovery. Current approaches, whether OCR-based pipelines

agentsarxiv-cs-cl
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

See Through the Noise: Improving Domain Generalization in Gaze Estimation

DGX agent

arXiv:2604.16562v1 Announce Type: new Abstract: Generalizable gaze estimation methods have garnered increasing attention due to their critical importance in real-world applications and have achieved s

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

Semantically Stable Image Composition Analysisvia Saliency and Gradient Vector Flow Fusion

DGX agent

arXiv:2604.16500v1 Announce Type: new Abstract: The reliable computational assessment of photographic composition requires features that are discriminative of spatial layout yet robust to semantic con

model-releasesarxiv-cs-cv
21 Apr 2026
Applications

SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe

DGX agent

arXiv:2410.05248v4 Announce Type: replace Abstract: To acquire instruction-following capabilities, large language models (LLMs) undergo instruction tuning, where they are trained on instruction-respon

applicationsarxiv-cs-cl
21 Apr 2026
Research

Singularity Formation: Synergy in Theoretical, Numerical and Machine Learning Approaches

DGX agent

arXiv:2604.16842v1 Announce Type: cross Abstract: This thesis develops numerical and theoretical approaches for understanding and analyzing singularity formation in Partial Differential Equations (PDE

researcharxiv-cs-lg
21 Apr 2026
Model Releases

SinkRouter: Sink-Aware Routing for Efficient Long-Context Decoding in Large Language and Multimodal Models

DGX agent

arXiv:2604.16883v1 Announce Type: new Abstract: In long-context decoding for LLMs and LMMs, attention becomes increasingly memory-bound because each decoding step must load a large amount of KV-cache

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Spatiotemporal Sycophancy: Negation-Based Gaslighting in Video Large Language Models

DGX agent

arXiv:2604.17873v1 Announce Type: new Abstract: Video Large Language Models (Vid-LLMs) have demonstrated remarkable performance in video understanding tasks, yet their robustness under conversational

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

SpeechMedAssist: Efficiently and Effectively Adapting Speech Language Models for Medical Consultation

DGX agent

arXiv:2601.04638v2 Announce Type: replace Abstract: Medical consultations are intrinsically speech-centric. However, most prior works focus on long-text-based interactions, which are cumbersome and pa

model-releasesarxiv-cs-cl
21 Apr 2026
Research

SpidR-Adapt: A Universal Speech Representation Model for Few-Shot Adaptation

DGX agent

arXiv:2512.21204v2 Announce Type: replace Abstract: Human infants, with only a few hundred hours of speech exposure, acquire basic units of new languages, highlighting a striking efficiency gap compar

researcharxiv-cs-cl
21 Apr 2026
Model Releases

SpiralFormer: Looped Transformers Can Learn Hierarchical Dependencies via Multi-Resolution Recursion

DGX agent

arXiv:2602.11698v2 Announce Type: replace Abstract: Recursive (looped) Transformers decouple computational depth from parameter depth by repeatedly applying shared layers, providing an explicit archit

model-releasesarxiv-cs-lg
21 Apr 2026
Research

SPOT: Single-Shot Positioning via Trainable Near-Field Rainbow Beamforming

DGX agent

arXiv:2511.11391v3 Announce Type: replace Abstract: Phase-time arrays, which integrate phase shifters (PSs) and true-time delays (TTDs), have emerged as a cost-effective architecture for generating fr

researcharxiv-cs-lg
21 Apr 2026
Model Releases

StepPO: Step-Aligned Policy Optimization for Agentic Reinforcement Learning

DGX agent

arXiv:2604.18401v1 Announce Type: new Abstract: General agents have given rise to phenomenal applications such as OpenClaw and Claude Code. As these agent systems (a.k.a. Harnesses) strive for bolder

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Support Sufficiency as Consequence-Sensitive Compression in Belief Arbitration

DGX agent

arXiv:2604.16434v1 Announce Type: cross Abstract: When a system commits to a hypothesis, much of the evidential structure behind that commitment is lost to compression. Standard accounts assume that s

safetyarxiv-cs-lg
21 Apr 2026
Research

Symmetry Guarantees Statistic Recovery in Variational Inference

DGX agent

arXiv:2604.18310v1 Announce Type: cross Abstract: Variational inference (VI) is a central tool in modern machine learning, used to approximate an intractable target density by optimising over a tracta

researcharxiv-cs-lg
21 Apr 2026
Safety

SynAgent: Generalizable Cooperative Humanoid Manipulation via Solo-to-Cooperative Agent Synergy

DGX agent

arXiv:2604.18557v1 Announce Type: new Abstract: Controllable cooperative humanoid manipulation is a fundamental yet challenging problem for embodied intelligence, due to severe data scarcity, complexi

safetyarxiv-cs-cv
21 Apr 2026
Agents

Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards

DGX agent

arXiv:2603.02008v2 Announce Type: replace Abstract: Effective exploration in reinforcement learning requires not only tracking where an agent has been, but also understanding how the agent perceives a

agentsarxiv-cs-lg
21 Apr 2026
Tutorials

This is what I’ve been cooking in the past 4 months . GPT Image 2 is over a massive 240 elo jump over the second place model, marking the bi…

DGX agent

This is what I’ve been cooking in the past 4 months . GPT Image 2 is over a massive 240 elo jump over the second place model, marking the biggest jump bigger than the rest of the leaderboard combined

tutorialsjeremy-howard--x
21 Apr 2026
Research

Time-Division Multiplexing Actuation in Tendon-Driven Arms: Lightweight Design and Fault Tolerance

DGX agent

arXiv:2604.16887v1 Announce Type: new Abstract: Robotic manipulators for aerospace applications require a delicate balance between lightweight construction and fault-tolerant operation to satisfy stri

researcharxiv-cs-ro
21 Apr 2026
Research

Topology Structure Optimization of Reservoirs Using GLMY Homology

DGX agent

arXiv:2509.11612v3 Announce Type: replace Abstract: Reservoir is an efficient network for time series processing. It is well known that network structure is one of the determinants of its performance.

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Towards a Data-Parameter Correspondence for LLMs: A Preliminary Discussion

DGX agent

arXiv:2604.17384v1 Announce Type: new Abstract: Large language model optimization has historically bifurcated into isolated data-centric and model-centric paradigms: the former manipulates involved sa

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Towards Generalizable Deepfake Image Detection with Vision Transformers

DGX agent

arXiv:2604.17376v1 Announce Type: new Abstract: In today's day and age, we face a challenge in detecting deepfake images because of the fast evolution of modern generative models and the poor generali

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Towards Initialization-dependent and Non-vacuous Generalization Bounds for Overparameterized Shallow Neural Networks

DGX agent

arXiv:2604.00505v2 Announce Type: replace Abstract: Overparameterized neural networks often show a benign overfitting property in the sense of achieving excellent generalization behavior despite the n

researcharxiv-cs-lg
21 Apr 2026
Research

Towards Reliable Testing of Machine Unlearning

DGX agent

arXiv:2604.16536v1 Announce Type: new Abstract: Machine learning components are now central to AI-infused software systems, from recommendations and code assistants to clinical decision support. As re

researcharxiv-cs-lg
21 Apr 2026
Safety

Towards Robust Text-to-Image Person Retrieval: Multi-View Reformulation for Semantic Compensation

DGX agent

arXiv:2604.18376v1 Announce Type: new Abstract: In text-to-image person retrieval tasks, the diversity of natural language expressions and the implicitness of visual semantics often lead to the proble

safetyarxiv-cs-cv
21 Apr 2026
Agents

Towards Self-Improving Error Diagnosis in Multi-Agent Systems

DGX agent

arXiv:2604.17658v1 Announce Type: cross Abstract: Large Language Model (LLM)-based Multi-Agent Systems (MAS) enable complex problem-solving but introduce significant debugging challenges, characterize

agentsarxiv-cs-cl
21 Apr 2026
Research

Trajectory-Restricted Optimization Conditions and Geometry-Aware Linear Convergence

DGX agent

arXiv:2604.17067v1 Announce Type: cross Abstract: Linear convergence of first-order methods is typically characterized by global optimization conditions whose constants reflect worst-case geometry of

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Triples and Knowledge-Infused Embeddings for Clustering and Classification of Scientific Documents

DGX agent

arXiv:2601.08841v2 Announce Type: replace Abstract: The increasing volume and complexity of scientific literature demand robust methods for organizing and understanding research documents. In this stu

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

TSVer: A Benchmark for Fact Verification Against Time-Series Evidence

DGX agent

arXiv:2511.01101v2 Announce Type: replace Abstract: Reasoning over temporal and numerical data, such as time series, is a crucial aspect of fact-checking. While many systems have recently been develop

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts

DGX agent

arXiv:2604.16542v1 Announce Type: cross Abstract: Safety guardrails have become an active area of research in AI safety, aimed at ensuring the appropriate behavior of large language models (LLMs). How

local-aiarxiv-cs-cl
21 Apr 2026
Tutorials

UGD: An Unsupervised Geometric Distance for Evaluating Real-world Noisy Point Cloud Denoising

DGX agent

arXiv:2604.16976v1 Announce Type: new Abstract: Point cloud denoising is a fundamental and crucial challenge in real-world point cloud applications. Existing quantitative evaluation metrics for point

tutorialsarxiv-cs-cv
21 Apr 2026
Model Releases

Unveiling Deepfakes: A Frequency-Aware Triple Branch Network for Deepfake Detection

DGX agent

arXiv:2604.17477v1 Announce Type: new Abstract: Advanced deepfake technologies are blurring the lines between real and fake, presenting both revolutionary opportunities and alarming threats. While it

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

User-Assistant Bias in LLMs

DGX agent

arXiv:2508.15815v3 Announce Type: replace Abstract: Modern large language models (LLMs) are typically trained and deployed using structured role tags (e.g. system, user, assistant, tool) that explicit

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Video Panels for Long Video Understanding

DGX agent

arXiv:2509.23724v2 Announce Type: replace Abstract: Recent Video-Language Models (VLMs) achieve promising results on long-video understanding, but their performance still lags behind that achieved on

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

VIDS: A Verified Imaging Dataset Standard for Medical AI

DGX agent

arXiv:2604.17525v1 Announce Type: cross Abstract: Medical imaging AI development is fundamentally dependent on annotated datasets, yet no existing standard provides machine-enforceable validation acro

model-releasesarxiv-cs-cv
21 Apr 2026
Applications

Visual-RRT: Finding Paths toward Visual-Goals via Differentiable Rendering

DGX agent

arXiv:2604.16388v1 Announce Type: cross Abstract: Rapidly-exploring random trees (RRTs) have been widely adopted for robot motion planning due to their robustness and theoretical guarantees. However,

applicationsarxiv-cs-cv
21 Apr 2026
Research

ViT^3: Unlocking Test-Time Training in Vision

DGX agent

arXiv:2512.01643v2 Announce Type: replace Abstract: Test-Time Training (TTT) has recently emerged as a promising direction for efficient sequence modeling. TTT reformulates attention operation as an o

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Voronoi-guided Bilateral 2D Gaussian Splatting for Arbitrary-Scale Hyperspectral Image Super-Resolution

DGX agent

arXiv:2604.17727v1 Announce Type: new Abstract: Most existing hyperspectral image super-resolution methods require modifications for different scales, limiting their flexibility in arbitrary-scale rec

model-releasesarxiv-cs-cv
21 Apr 2026
Agents

We are entering an extremely exciting era for open-weight models. Kimi K2.6 now feels like a top agentic model. I took it for a spin via @Fi…

DGX agent

We are entering an extremely exciting era for open-weight models. Kimi K2.6 now feels like a top agentic model. I took it for a spin via @FireworksAI_HQ fast inference APIs. Kimi K2.6 has impressive a

agentsdair-ai--x
21 Apr 2026
Safety

Weakly-Supervised Referring Video Object Segmentation through Text Supervision

DGX agent

arXiv:2604.17797v1 Announce Type: new Abstract: Referring video object segmentation (RVOS) aims to segment the target instance in a video, referred by a text expression. Conventional approaches are mo

safetyarxiv-cs-cv
21 Apr 2026
Research

What If Consensus Lies? Selective-Complementary Reinforcement Learning at Test Time

DGX agent

arXiv:2603.19880v2 Announce Type: replace Abstract: Test-Time Reinforcement Learning (TTRL) enables Large Language Models (LLMs) to enhance reasoning capabilities on unlabeled test streams by deriving

researcharxiv-cs-lg
21 Apr 2026
Research

What makes an entity salient in discourse?

DGX agent

arXiv:2508.16464v2 Announce Type: replace Abstract: Entities in discourse vary in salience: main participants, objects and locations stay prominent, while others are quickly forgotten, raising questio

researcharxiv-cs-cl
21 Apr 2026
Model Releases

When Earth Foundation Models Meet Diffusion: An Application to Land Surface Temperature Super-Resolution

DGX agent

arXiv:2604.16841v1 Announce Type: new Abstract: Land surface temperature (LST) super-resolution is important for environmental monitoring. However, it remains challenging as coarse thermal observation

model-releasesarxiv-cs-cv
21 Apr 2026
Research

When Misinformation Speaks and Converses: Rethinking Fact-Checking in Audio Platforms

DGX agent

arXiv:2604.16767v1 Announce Type: new Abstract: Audio platforms have evolved beyond entertainment. They have become central to public discourse, from podcasts and radio to WhatsApp voice notes and liv

researcharxiv-cs-cl
21 Apr 2026
Safety

Why Agents Compromise Safety Under Pressure

DGX agent

arXiv:2603.14975v2 Announce Type: replace-cross Abstract: Large Language Model agents deployed in complex environments frequently encounter a conflict between maximizing goal achievement and adhering

safetyarxiv-cs-cl
21 Apr 2026
Safety

Why Low-Precision Transformer Training Fails: An Analysis on Flash Attention

DGX agent

arXiv:2510.04212v3 Announce Type: replace Abstract: The pursuit of computational efficiency has driven the adoption of low-precision formats for training transformer models. However, this progress is

safetyarxiv-cs-lg
21 Apr 2026
Research

Why Training-Free Token Reduction Collapses: The Inherent Instability of Pairwise Scoring Signals

DGX agent

arXiv:2604.16745v1 Announce Type: cross Abstract: Training-free token reduction methods for Vision Transformers (ToMe, ToFu, PiToMe, and MCTF) employ different scoring mechanisms, yet they share a clo

researcharxiv-cs-cv
21 Apr 2026
Applications

A Q-learning-based QoS-aware multipath routing protocol in IoMT-based wireless body area network

DGX agent

arXiv:2604.15489v1 Announce Type: cross Abstract: The Internet of Medical Things (IoMT) enables intelligent healthcare services but faces challenges such as dynamic topology, energy constraints, and d

applicationsarxiv-cs-ai
20 Apr 2026
Safety

“AI is at best a functional mimic, not a conscious experiencing subject. …. The real moral issue lies not in making AI conscious …. but in a…

DGX agent

“AI is at best a functional mimic, not a conscious experiencing subject. …. The real moral issue lies not in making AI conscious …. but in avoiding transforming humans into zombies” @GaryMarcus @OEIAC

safetygary-marcus--x
20 Apr 2026
← Previous
1…232233234235236…254
Next →