AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,603 results
Model Releases

CAOA -- Completion-Assisted Object-CAD Alignment

DGX agent

arXiv:2606.18429v2 Announce Type: replace Abstract: Accurately aligning CAD models to their corresponding objects in indoor RGB-D scans is a central challenge in 3D semantic reconstruction. The task r

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

CapRiCorn-1K: A Comprehensive Benchmark for Video Captioning and Subject Referential Consistency Across Temporal Scales

DGX agent

arXiv:2606.21949v1 Announce Type: new Abstract: Accurate and comprehensive video captions with consistent subject references are critical for downstream understanding and generation tasks. However, fe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Catching Lies Without Sending the Video: Privacy-Preserving Multimodal Deception Detection

DGX agent

arXiv:2606.22699v1 Announce Type: new Abstract: Frontier multimodal models can guess whether a person is lying from a testimony video. To do so, they stream that raw face and voice to a third-party mo

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CDER-SME: A Cross-Device Event-RGB Micro-Expression Dataset under Multi-Level Stress Induction

DGX agent

arXiv:2606.20715v1 Announce Type: new Abstract: Micro-expression recognition (MER) in realistic scenarios demands high temporal sensitivity and ecological validity, yet existing benchmarks are largely

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CFAgentBench: A Reproducible Environment and Benchmark for Autonomous Construction-Finance Agents

DGX agent

arXiv:2606.22000v1 Announce Type: cross Abstract: We introduce CFAgentBench, a reproducible, self-hostable environment and benchmark for autonomous construction-finance agents: a CFO/controller-class

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Chains That See, Answers That Don't: A Multi-Aspect Evaluation Recipe for Forced Chain-of-Thought on Video-MME

DGX agent

arXiv:2606.22862v1 Announce Type: new Abstract: Forced chain-of-thought (CoT) is widely assumed to make vision-language models more reliable on video question answering. We propose a small three-probe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Chehre: An Emoji-Prompted Video Dataset for Perceptually Diverse Facial Expression Recognition

DGX agent

arXiv:2606.21657v1 Announce Type: new Abstract: Facial expressions are nonverbal social signals used in human interaction, but facial expression recognition datasets often focus on static images, basi

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Chem2Gen-Bench: Benchmarking Chemical-to-Genetic Translation in Perturbation Response Space

DGX agent

arXiv:2606.21109v1 Announce Type: new Abstract: Virtual-cell and perturbation models are increasingly used to predict cellular responses for biomedical discovery, but chemical and genetic perturbation

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

CheXpercept: A Benchmark for Evaluating Expert-Level Lesion Perception in Chest X-rays

DGX agent

arXiv:2606.21020v1 Announce Type: new Abstract: The evaluation of vision-language models (VLMs) for chest X-ray (CXR) analysis has largely been limited to disease-presence classification without visua

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

ChronoLock: Protecting Videos from Unauthorized Text-to-Video Personalization

DGX agent

arXiv:2606.21146v1 Announce Type: new Abstract: Text-to-video (T2V) diffusion models have made it increasingly easy to synthesize realistic and temporally coherent videos, while recent personalization

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Claude is really proactive with Claude Tag. You don’t need to prompt it to do work, it can do work proactively based on your instructions. I…

DGX agent

Claude is really proactive with Claude Tag. You don’t need to prompt it to do work, it can do work proactively based on your instructions. It has excellent memory and access to your data, so it can be

model-releasesboris-cherny--x
23 Jun 2026
Model Releases

Claude Tag is an incredible new form factor for agents, so I think it's going to take some time to figure out the best practices, but these …

DGX agent

Claude Tag is an incredible new form factor for agents, so I think it's going to take some time to figure out the best practices, but these are some of my favorites 🧵 Introducing Claude Tag, a new way

model-releasesthariq--x
23 Jun 2026
Model Releases

Clipping the Price of Adaptivity at the Tail

DGX agent

arXiv:2606.22669v1 Announce Type: new Abstract: Adaptive stochastic convex optimization (SCO) methods face a fundamental ``price of adaptivity'' barrier: under the standard set of assumptions, they ca

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Cluster-Specific Localized Drift Detection for Efficient Batch Model Adaptation under Controlled Distribution Shift

DGX agent

arXiv:2606.22026v1 Announce Type: new Abstract: Machine learning systems deployed in dynamic environments frequently operate under nonstationary data distributions, where controlled distribution shift

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

CodePercept: Code-Grounded Visual STEM Perception for MLLMs

DGX agent

arXiv:2603.10757v2 Announce Type: replace Abstract: When MLLMs fail at Science, Technology, Engineering, and Mathematics (STEM) visual reasoning, a fundamental question arises: is it due to perceptual

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Comparative Evaluation of Machine Learning and Deep Learning Models for Wound-Rotor Synchronous Motor Performance Prediction

DGX agent

arXiv:2606.21230v1 Announce Type: new Abstract: Wound rotor synchronous motors have emerged as a strong alternative that eliminates dependence on REEs. However, WRSM design requires the simultaneous o

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Concept-Constrained Prompt Learning for Few-Shot CLIP Adaptation

DGX agent

arXiv:2606.22567v1 Announce Type: new Abstract: Few-shot prompt learning is an effective strategy for adapting CLIP to downstream tasks, but class-only prompt optimization can overfit base-class super

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Confidently Wrong: Severity-Aware Calibration of Prompt-Injection Detectors under Attack Shift

DGX agent

arXiv:2606.22659v1 Announce Type: cross Abstract: Prompt-injection detectors are deployed as guards: a model scores an input and a downstream system trusts or blocks it on that score. I study the conf

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

ConnectomeBench2: A Unified Benchmark for Automated Connectomic Proofreading

DGX agent

arXiv:2606.21116v1 Announce Type: new Abstract: Proofreading--correcting segmentation errors in 3D brain reconstructions--is the rate-limiting step in synapse-resolution connectomics. We release Conne

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CoRDE: Concept-Prior Routed Diffusion Experts for Structural Generalization in Robot Manipulation

DGX agent

arXiv:2606.21935v1 Announce Type: new Abstract: Diffusion models excel at capturing multi-modal action distributions in robot imitation learning. However, in multi-task and long-horizon scenarios, mon

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

CoSA: Correlation-Guided Change Attention with Learnable Residual Gating for Remote Sensing Change Detection

DGX agent

arXiv:2606.21932v1 Announce Type: new Abstract: Remote sensing change detection (CD) from bi-temporal imagery is critical for applications such as urban monitoring, disaster assessment, and environmen

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CourseBlueprint: A Structured Pipeline for Adaptive Pedagogical Video Generation Grounded in Course Corpora

DGX agent

arXiv:2606.20608v1 Announce Type: cross Abstract: Generative text-to-video systems can produce visually fluent educational clips, but they rarely encode the pedagogical content knowledge (PCK) needed

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CRAX: Fast Safe Reinforcement Learning Benchmarking

DGX agent

arXiv:2606.20376v2 Announce Type: replace Abstract: Safety is a core concern for deploying reinforcement learning (RL) agents in real-world domains such as robotics and autonomous driving. While bench

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model

DGX agent

arXiv:2606.22317v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) is widely viewed as a promising path toward continuously improving large language models. Recent w

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Curvature-aware 3D length estimation of greenhouse cucumbers using RGB-D imaging and cubic spline arc-length integration

DGX agent

arXiv:2606.22439v1 Announce Type: new Abstract: Commercial greenhouse cucumber production is graded by fruit length, which drives harvest scheduling, labour allocation, and logistics. Manual measureme

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

CVSBench: A Comprehensive Benchmark for Cross-view Spatial Reasoning and Dreaming

DGX agent

arXiv:2606.22476v1 Announce Type: new Abstract: Humans can effortlessly reason about scenes across different viewpoints, yet it remains unclear whether Vision-Language Models (VLMs) possess similar cr

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Data Pruning: Redundant, Problematic, and Interdependent Samples

DGX agent

arXiv:2606.21916v1 Announce Type: new Abstract: The performance of deep learning models is affected by not only data quantity but also data quality. Data pruning is a process by which practitioners ca

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams

DGX agent

arXiv:2606.21337v1 Announce Type: new Abstract: Massive unstructured multimodal streams suffer from high 'data entropy,' impeding both efficient human knowledge acquisition and high-quality AI post-tr

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Decoupling the Declarative from the Procedural in Vision-Language-Action Models

DGX agent

arXiv:2606.21496v1 Announce Type: cross Abstract: Deploying generalist robotic agents in the real world requires transferable skills. Specifically, a policy trained to clone a behavior from object-spe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Deep Learning-Based Sign Language Recognition from Videos and Cross-Lingual Translation to Indian Vernaculars

DGX agent

arXiv:2606.22494v1 Announce Type: cross Abstract: Sign language is a primary mode of communication for the global deaf and hard-of-hearing community, yet automated tools that recognize sign gestures f

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Demystifying Numerical Instability in LLM Inference: Achieving Reproducible Inference for Mission-Critical Tasks with HEAL

DGX agent

arXiv:2606.21023v1 Announce Type: new Abstract: As Large Language Models (LLMs) deploy into mission-critical domains (e.g., finance, medicine, and law), output reproducibility has become a strict syst

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Deploy from Claude Design to Vercel

DGX agent

This Vercel changelog entry describes integration between Claude Design and Vercel that enables direct deployment of designs to the Vercel platform. The feature likely streamlines the workflow for dev

model-releasesvercel-blog
23 Jun 2026
Model Releases

Detail++: Training-Free Detail Enhancer for T2I Diffusion Models

DGX agent

arXiv:2507.17853v3 Announce Type: replace Abstract: Recent advances in text-to-image (T2I) generation have led to impressive visual results. However, these models still face significant challenges whe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Disentangling Intrinsic Importance from Emergent Structure in Multi-Expert Orchestration

DGX agent

arXiv:2602.04291v2 Announce Type: replace Abstract: Multi-expert systems, where multiple Large Language Models (LLMs) collaborate to solve complex tasks, are increasingly adopted for high-performance

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Distilling Collaborative Dynamics into Latent Space for Implicit Coordination in Decentralized Multi-Agent Manipulation

DGX agent

arXiv:2606.22982v1 Announce Type: new Abstract: Multi-arm manipulation demands precise spatiotemporal coordination, yet many centralized approaches scale poorly as team size increases. To address this

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

Distinguishing indistinguishable attractors: Unsupervised anomaly detection with reservoir computers

DGX agent

arXiv:2606.21322v1 Announce Type: new Abstract: Detecting when a nonlinear dynamical system departs from its normal regime is a recurring problem across the sciences, from cardiology to climate and en

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Distributed Model Predictive Control with Adaptive Safety Zones for Multi-Fleet Drone Operations

DGX agent

arXiv:2606.20651v1 Announce Type: cross Abstract: Autonomous drone swarms in space-constrained environments such as warehouses, inspection corridors, and urban delivery routes must share limited airsp

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

Distribution-Aware Robust Bilevel Optimization: Quantile-Guided Huber Updates in Two-Timescale Stochastic Approximation

DGX agent

arXiv:2606.22436v1 Announce Type: new Abstract: Bilevel optimization (BLO) is fundamental to hierarchical decision-making but suffers from critical instability under heavy-tailed stochastic noise. Exi

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Do Location Encoders Capture Spatial Effects? A GeoShapley Benchmark Across Scales

DGX agent

arXiv:2606.23453v1 Announce Type: new Abstract: Location encoders transform geographic coordinates into high dimensional embeddings for downstream machine learning, but it is unclear how well these re

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Do Modern Video-LLMs Need to Listen? A Benchmark Audit and Scalable Remedy

DGX agent

arXiv:2509.17901v4 Announce Type: replace Abstract: Speech and audio encoders developed over years of community effort are routinely excluded from video understanding pipelines, not because they fail,

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Does RoPE Prevent or Degrade Retrieval Heads? A Mechanistic Analysis Across Model Families

DGX agent

arXiv:2606.21249v1 Announce Type: new Abstract: Retrieval heads, attention heads that copy information from earlier context to the current position, have been proposed as the mechanistic substrate for

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Double-Diffusion: Balancing Speed, Accuracy, and Uncertainty in Probabilistic Forecasting for Urban Sensor Networks

DGX agent

arXiv:2506.23053v3 Announce Type: replace Abstract: Urban sensor networks need forecasts that are accurate, carry useful uncertainty, and refresh fast enough to act on as new readings arrive. These go

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

DR-Mamba: Automatic Inference-Time Domain Adaptation for Document Image Binarization via Sample-Conditioned Detail-Background Suppression

DGX agent

arXiv:2606.22625v1 Announce Type: new Abstract: Degraded document image binarization is sensitive to domain shifts caused by paper aging, bleed-through, stains, shadows, and uneven illumination, and t

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

DrivingVoxels: Compositional Sparse Voxel Rasterization for Dynamic Driving Scene Reconstruction

DGX agent

arXiv:2606.23031v1 Announce Type: new Abstract: Reconstructing dynamic urban scenes remains challenging due to the unbounded nature of driving environments and the presence of multiple dynamic objects

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

DrugBench: Evaluating AI Control Protocols for Medication Harm Mitigation

DGX agent

arXiv:2606.20663v1 Announce Type: cross Abstract: Large Language Models have the potential to expand and improve the access to clinical information by enabling new ways of interacting with medical kno

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Duet: Dual-Robot Understanding via Efficient Teaching

DGX agent

arXiv:2606.20990v1 Announce Type: new Abstract: Dual-robot collaboration enables tasks that exceed the reach and payload of a single robot, such as collaboratively transporting objects across environm

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

Dynamics, stability, and energy efficiency of an energy-recycling rimless wheel with spring-clutch legs

DGX agent

arXiv:2606.22073v1 Announce Type: new Abstract: This paper proposes an energy-recycling rimless wheel with spring-clutch legs. The proposed mechanism uses a lockable clutch to store part of the impact

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation

DGX agent

arXiv:2606.22918v1 Announce Type: new Abstract: Maintaining physical consistency in video generators and world models increasingly relies on vision-language models (VLMs) as automated judges that prov

model-releasesarxiv-cs-cv
23 Jun 2026
← Previous
1…174175176177178…471
Next →