AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Applications

AVI-HT: Adaptive Vision-IMU Fusion for 3D Hand Tracking

DGX agent

arXiv:2605.21714v1 Announce Type: new Abstract: We present AVI-HT, an adaptive visual-IMU fusion approach for tracking 3D hand poses by jointly modeling the egocentric image with on-glove 6-DoF IMU si

applicationsarxiv-cs-cv
22 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ChronoMedKG: A Temporally-Grounded Biomedical Knowledge Graph and Benchmark for Clinical Reasoning

DGX agent

arXiv:2605.22734v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) treat disease associations as static facts, but temporal information is crucial for clinical reasoning, e.g., a sympto

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

COCOTree: A Dataset and Benchmark for Open Tree-Structured Visual Decomposition

DGX agent

arXiv:2605.22068v1 Announce Type: new Abstract: We formalize and enable the task of open tree decomposition, which segments an image into hierarchical trees of visual components with unconstrained gra

model-releasesarxiv-cs-cv
22 May 2026
Applications

Codec-Robust Attacks on Audio LLMs

DGX agent

arXiv:2605.20519v1 Announce Type: cross Abstract: Prior attacks on Audio Large Language Models (Audio LLMs) demonstrated that carefully crafted waveform-domain perturbations can force targeted adversa

applicationsarxiv-cs-ai
22 May 2026
Model Releases

Cohesion-6K: An Arabic Dataset for Analyzing Social Cohesion and Conflict in Online Discourse

DGX agent

arXiv:2605.22447v1 Announce Type: new Abstract: The study of online discourse has become central to understanding societal polarization. While much research has focused on detecting overt toxicity, th

model-releasesarxiv-cs-cl
22 May 2026
Research

D3Seg: Dependency-Aware Diffusion for Brain Tumor Segmentation with Missing Modalities

DGX agent

arXiv:2605.22249v1 Announce Type: new Abstract: Accurate brain tumor segmentation using multiparametric MRI is critical for effective treatment planning. However, in clinical settings, complete acquis

researcharxiv-cs-cv
22 May 2026
Research

Entropy-Guided Self-Supervised Learning for Medical Image Classification

DGX agent

arXiv:2605.21970v1 Announce Type: cross Abstract: Accurate and robust medical image classification is paramount for early disease diagnosis and treatment planning. However, challenges such as limited

researcharxiv-cs-cv
22 May 2026
Model Releases

FashionLens: Toward Versatile Fashion Image Retrieval via Task-Adaptive Learning

DGX agent

arXiv:2605.22552v1 Announce Type: new Abstract: Fashion image retrieval is a cornerstone of modern e-commerce systems. A unified framework that supports diverse query formats and search intentions is

model-releasesarxiv-cs-cv
22 May 2026
Research

Geometry-Adaptive Explainer for Faithful Dictionary-Based Interpretability under Distribution Shift

DGX agent

arXiv:2605.21849v1 Announce Type: cross Abstract: Mechanistic interpretability aims to explain a model's behavior by identifying causally responsible internal structures. Dictionary-based explainers s

researcharxiv-cs-cl
22 May 2026
Model Releases

ha ha, so much for step change? maybe this problem was just easier than some?

DGX agent

ha ha, so much for step change? maybe this problem was just easier than some? The standard GPT-5.5 reproduced the proof ~ 👇 https://chatgpt.com/share/6a0e9e04-8cb0-8332-a4f1-ec68acd2e03e You don't nee

model-releasesgary-marcus--x
22 May 2026
Model Releases

Highlighting the new WebGPU backend in llama.cpp/ggml The work to bring full-fledged WebGPU support in llama.cpp started about an year and a…

DGX agent

Highlighting the new WebGPU backend in llama.cpp/ggml The work to bring full-fledged WebGPU support in llama.cpp started about an year and a half ago. It has been lead by @reeselevine and team at USCS

model-releasesgeorgi-gerganov--x
22 May 2026
Model Releases

HyLoVQA: Dynamic Hypernetwork-Generated Low-Rank Adaptation for Continual Visual Question Answering

DGX agent

arXiv:2605.22035v1 Announce Type: cross Abstract: Continual Visual Question Answering (VQA) requires learning from non-stationary streams of visual inputs and questions while preserving past knowledge

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

InfVSR: Breaking Length Limits of Generic Video Super-Resolution

DGX agent

arXiv:2510.00948v2 Announce Type: replace Abstract: Real-world videos often extend over thousands of frames. Existing generative video super-resolution (VSR) approaches, however, face two persistent c

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Link to GPT 5.5 on the recent Erdo problem: https://x.com/maxiao54704/status/2057484153755480537?s=61

DGX agent

Link to GPT 5.5 on the recent Erdo problem: https://x.com/maxiao54704/status/2057484153755480537?s=61 The standard GPT-5.5 reproduced the proof ~ 👇 https://chatgpt.com/share/6a0e9e04-8cb0-8332-a4f1-ec

model-releasesgary-marcus--x
22 May 2026
Model Releases

MLLMs Know When Before Speaking: Revealing and Recovering Temporal Grounding via Attention Cues

DGX agent

arXiv:2605.21954v1 Announce Type: new Abstract: Video temporal grounding (VTG), which localizes the start and end times of a queried event in an untrimmed video, is a key test of whether multimodal la

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

MM-Conv: A Multimodal Dataset and Benchmark for Context-Aware Grounding in 3D Dialogue

DGX agent

arXiv:2605.21796v1 Announce Type: cross Abstract: Grounding language in the physical world requires AI systems to interpret references that emerge dynamically during conversation. While current vision

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

MOTOR: A Multimodal Dataset for Two-Wheeler Rider Behavior Understanding

DGX agent

arXiv:2605.22550v1 Announce Type: new Abstract: Two-wheelers account for a disproportionately high share of road fatalities in the Global South. Research on two-wheeler rider behavior, however, lags f

model-releasesarxiv-cs-cv
22 May 2026
Safety

NaviAgent: Graph-Driven Bilevel Planning for Scalable Tool Orchestration

DGX agent

arXiv:2506.19500v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) increasingly act as function-call agents that invoke external tools to tackle tasks beyond their static knowledge

safetyarxiv-cs-cl
22 May 2026
Safety

Open-source LLMs administer maximum electric shocks in a Milgram-like obedience experiment

DGX agent

arXiv:2605.21401v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that make sequences of decisions over extended interactions in high-stakes

safetyarxiv-cs-ai
22 May 2026
Model Releases

PromptNCE: Pointwise Mutual Information Predictions Using Only LLMs and Contrastive Estimation Prompts

DGX agent

arXiv:2605.21776v1 Announce Type: new Abstract: Estimating mutual information from text usually requires training a task-specific critic, which limits its use in low-data settings. We ask whether larg

model-releasesarxiv-cs-cl
22 May 2026
Applications

Quantizing Whisper-small: How design choices affect ASR performance

DGX agent

arXiv:2511.08093v2 Announce Type: replace-cross Abstract: Large speech recognition models like Whisper-small achieve high accuracy but are difficult to deploy on edge devices due to their high computa

applicationsarxiv-cs-cl
22 May 2026
Local Ai

SegCompass: Exploring Interpretable Alignment with Sparse Autoencoders for Enhanced Reasoning Segmentation

DGX agent

arXiv:2605.22658v1 Announce Type: new Abstract: While large language models provide strong compositional reasoning, existing reasoning segmentation pipelines fail to transparently connect this reasoni

local-aiarxiv-cs-cv
22 May 2026
Model Releases

SFN-YOLO: Towards Free-Range Poultry Detection via Scale-aware Fusion Networks

DGX agent

arXiv:2509.17086v2 Announce Type: replace Abstract: Detecting and localizing poultry is essential for advancing smart poultry farming. Despite the progress of detection-centric methods, challenges per

model-releasesarxiv-cs-cv
22 May 2026
Agents

SpecHop: Continuous Speculation for Accelerating Multi-Hop Retrieval Agents

DGX agent

arXiv:2605.21965v1 Announce Type: new Abstract: Large language models increasingly use external tools such as web search and document retrieval to solve information-intensive tasks. However, multi-hop

agentsarxiv-cs-cl
22 May 2026
Research

ST-SimDiff: Balancing Spatiotemporal Similarity and Difference for Efficient Video Understanding with MLLMs

DGX agent

arXiv:2605.22158v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) face significant computational overhead when processing long videos due to the massive number of visual token

researcharxiv-cs-cv
22 May 2026
Model Releases

Steins;Gate Drive: Semantic Safety Arbitration over Structured Futures for Latency-Decoupled LLM Planning

DGX agent

arXiv:2605.22456v1 Announce Type: new Abstract: Cloud-hosted LLM driver agents provide useful semantic judgments, but their inference latency exceeds stepwise vehicle-control windows. Learned world mo

model-releasesarxiv-cs-ro
22 May 2026
Model Releases

SURGE: An Event-Centric Social Media Sentiment Time Series Benchmark with Interaction Structure

DGX agent

arXiv:2605.21198v1 Announce Type: cross Abstract: Public events on social media generate large volumes of discussion whose collective dynamics carry direct value for opinion forecasting and crisis res

model-releasesarxiv-cs-ai
22 May 2026
Local Ai

Training low resolution then switching to high resolution later?

DGX agent

A discussion of progressive training approaches for diffusion models where training begins at lower scale factors and progressively increases to target scale factors, leveraging previously trained mod

local-air-stablediffusion
22 May 2026
Model Releases

Training-Trajectory-Aware Token Selection

DGX agent

arXiv:2601.10348v2 Announce Type: replace Abstract: Efficient distillation is a key pathway for converting expensive reasoning capability into deployable efficiency, yet in the frontier regime where t

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation

DGX agent

arXiv:2605.22355v1 Announce Type: new Abstract: Public transit route planning traditionally depends on structured map infrastructure and complex routing engines, and no existing dataset supports train

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Two is better than one: A Collapse-free Multi-Reward RLIF Training Framework

DGX agent

arXiv:2605.22620v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning ability of LLMs, but often depends on external supervis

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

UniVL: Unified Vision-Language Embedding for Spatially Grounded Contextual Image Generation

DGX agent

arXiv:2605.21611v1 Announce Type: new Abstract: We introduce spatially grounded contextual image generation, a controllable image generation task that reframes the conditioning paradigm. Instead of su

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Another win for open-source robotics! 🔥 @huggingface just released a fully open-source humanoid robot, and you can build one for $2,500. I'…

DGX agent

Another win for open-source robotics! 🔥 @huggingface just released a fully open-source humanoid robot, and you can build one for $2,500. I'm a huge advocate of open-source in robotics space. Why? Robo

model-releasesclem-delangue--x
21 May 2026
Model Releases

APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents

DGX agent

arXiv:2605.21240v1 Announce Type: new Abstract: LLM agents have shown strong performance across a wide range of complex tasks, including interactive environments that require long-horizon decision mak

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

APM: Evaluating Style Personalization in LLMs with Arbitrary Preference Mappings

DGX agent

arXiv:2605.21063v1 Announce Type: new Abstract: Typical LLM responses tend to follow a default style, even though users often have distinct preferences regarding tone, verbosity, and formality that th

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Batched Single-Index Global Multi-Armed Bandits with Covariates

DGX agent

arXiv:2503.00565v3 Announce Type: replace-cross Abstract: The multi-armed bandits (MAB) framework is a widely used approach for sequential decision-making, where a decision-maker selects an arm in eac

model-releasesarxiv-cs-lg
21 May 2026
Research

Beyond Routing: Characterising Expert Tuning and Representation in Vision Mixture-of-Experts

DGX agent

arXiv:2605.20610v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models are often interpreted by analysing which categories are routed to which experts. However, routing alone does not reveal

researcharxiv-cs-cv
21 May 2026
Research

Block-Sparse Global Attention for Efficient Multi-View Geometry Transformers

DGX agent

arXiv:2509.07120v2 Announce Type: replace Abstract: Efficient and accurate feed-forward multi-view reconstruction has long been an important task in computer vision. Recent transformer-based models li

researcharxiv-cs-cv
21 May 2026
Safety

Choose Wisely and Privately: Proactive Client Selection for Fair and Efficient Federated Learning

DGX agent

arXiv:2605.20975v1 Announce Type: new Abstract: Federated Learning enables collaborative model training across decentralized data sources without data transfer. Averaging-based FL is limited by the pr

safetyarxiv-cs-lg
21 May 2026
Model Releases

ChunkFT: Byte-Streamed Optimization for Memory-Efficient Full Fine-Tuning

DGX agent

arXiv:2605.21177v1 Announce Type: cross Abstract: This work presents extsc{ChunkFT}, a memory-efficient fine-tuning framework that reformulates full-parameter fine-tuning around a dynamically activate

model-releasesarxiv-cs-cl
21 May 2026
Agents

Code Generation by Differential Test Time Scaling

DGX agent

arXiv:2605.20473v1 Announce Type: cross Abstract: Test-time scaling has emerged as a promising approach for improving code generation by exploring large solution spaces at inference time. However, exi

agentsarxiv-cs-lg
21 May 2026
Safety

Cross-lingual robustness of LLM-brain alignment and its computational roots

DGX agent

arXiv:2605.21049v1 Announce Type: new Abstract: Large language models (LLMs) reliably predict neural activity during language comprehension and transformer depth has been interpreted as mirroring hier

safetyarxiv-cs-cl
21 May 2026
Safety

Cumulative Meta-Learning from Active Learning Queries for Robustness to Spurious Correlations

DGX agent

arXiv:2605.20771v1 Announce Type: new Abstract: Spurious correlations in real-world datasets cause machine learning models to rely on irrelevant patterns, undermining reliability, generalization, and

safetyarxiv-cs-lg
21 May 2026
Model Releases

Deltaynamics: Language-Based Representation for Inferring Rigid-Body Dynamics From Videos

DGX agent

arXiv:2605.20576v1 Announce Type: new Abstract: Inferring rigid-body physical states and properties from monocular videos is a fundamental step toward physics-based perception and simulation. Existing

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Diagnosing Overhead in Dispatch Operations: Cross-architecture Observatory

DGX agent

arXiv:2605.20982v1 Announce Type: cross Abstract: AlltoAll dispatch is the dominant bottleneck of MoE expert parallelism, and the interconnect community has responded with four families of mitigations

model-releasesarxiv-cs-lg
21 May 2026
Agents

Draw2Think: Harnessing Geometry Reasoning through Constraint Engine Interaction

DGX agent

arXiv:2605.20743v1 Announce Type: cross Abstract: Vision-language models solve geometry problems with rising accuracy, yet their intermediate states remain latent and unverifiable: a relation expresse

agentsarxiv-cs-cl
21 May 2026
Model Releases

DySink: Dynamic Frame Sinks for Autoregressive Long Video Generation

DGX agent

arXiv:2605.21028v1 Announce Type: new Abstract: Autoregressive long video generation often adopts bounded-memory streaming for efficiency, typically combining local windows for short-term continuity w

model-releasesarxiv-cs-cv
21 May 2026
Research

Efficient Table QA via TableGrid Navigation and Progressive Inference Prompting

DGX agent

arXiv:2605.20254v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promising results on NLP tasks, however, their performance on tabular data still needs research attention, bec

researcharxiv-cs-cv
21 May 2026
← Previous
1…715716717718719…1371
Next →