AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,332 results
Model Releases

Detecting LLM-Generated Spam Reviews by Integrating Language Model Embeddings and Graph Neural Network

DGX agent

arXiv:2510.01801v2 Announce Type: replace Abstract: The rise of large language models (LLMs) has enabled the generation of highly persuasive spam reviews that closely mimic human writing. These review

model-releasesarxiv-cs-cl
21 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

DifFoundMAD: Foundation Models meet Differential Morphing Attack Detection

DGX agent

arXiv:2604.17961v1 Announce Type: new Abstract: In this work, we introduce DifFoundMAD, a parameter-efficient D-MAD framework that exploits the generalisation capabilities of vision foundation models

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining

DGX agent

arXiv:2604.16391v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have shown great potential in building generalist robots, but still face a dilemma-misalignment of 2D image foreca

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Do LLM-derived graph priors improve multi-agent coordination?

DGX agent

arXiv:2604.17191v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) is crucial for AI systems that operate collaboratively in distributed and adversarial settings, particularly i

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

DocQAC: Adaptive Trie-Guided Decoding for Effective In-Document Query Auto-Completion

DGX agent

arXiv:2604.18257v1 Announce Type: cross Abstract: Query auto-completion (QAC) has been widely studied in the context of web search, yet remains underexplored for in-document search, which we term DocQ

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Document-as-Image Representations Fall Short for Scientific Retrieval

DGX agent

arXiv:2604.18508v1 Announce Type: cross Abstract: Many recent document embedding models are trained on document-as-image representations, embedding rendered pages as images rather than the underlying

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Domain-oriented RAG Assessment (DoRA): Synthetic Benchmarking for RAG-based Question Answering on Defense Documents

DGX agent

arXiv:2604.17943v1 Announce Type: new Abstract: Open-domain RAG benchmarks over public corpora can overestimate deployment performance due to pretraining overlap and weak attribution requirements. We

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

DORA Explorer: Improving the Exploration Ability of LLMs Without Training

DGX agent

arXiv:2604.17244v1 Announce Type: new Abstract: Despite the rapid progress, LLMs for sequential decision-making (i.e., LLM agents) still struggle to produce diverse outputs. This leads to insufficient

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

DREAM: Dynamic Retinal Enhancement with Adaptive Multi-modal Fusion for Expert Precision Medical Report Generation

DGX agent

arXiv:2604.17209v1 Announce Type: new Abstract: Automating medical reports for retinal images requires a sophisticated blend of visual pattern recognition and deep clinical knowledge. Current Large Vi

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking

DGX agent

arXiv:2507.20879v3 Announce Type: replace Abstract: The advent of Vision-Language Models (VLMs) has significantly advanced end-to-end autonomous driving, demonstrating powerful reasoning abilities for

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

DSH-Bench: A Difficulty- and Scenario-Aware Benchmark with Hierarchical Subject Taxonomy for Subject-Driven Text-to-Image Generation

DGX agent

arXiv:2603.08090v2 Announce Type: replace Abstract: Significant progress has been achieved in subject-driven text-to-image (T2I) generation, which aims to synthesize new images depicting target subjec

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Dual-stream Spatio-Temporal GCN-Transformer Network for 3D Human Pose Estimation

DGX agent

arXiv:2604.17688v1 Announce Type: new Abstract: 3D human pose estimation is a classic and important research direction in the field of computer vision. In recent years, Transformer-based methods have

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

DuConTE: Dual-Granularity Text Encoder with Topology-Constrained Attention for Text-attributed Graphs

DGX agent

arXiv:2604.17411v1 Announce Type: new Abstract: Text-attributed graphs integrate semantic information of node texts with topological structure, offering significant value in various applications such

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

DuQuant++: Fine-grained Rotation Enhances Microscaling FP4 Quantization

DGX agent

arXiv:2604.17789v1 Announce Type: cross Abstract: The MXFP4 microscaling format, which partitions tensors into blocks of 32 elements sharing an E8M0 scaling factor, has emerged as a promising substrat

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

E3VS-Bench: A Benchmark for Viewpoint-Dependent Active Perception in 3D Gaussian Splatting Scenes

DGX agent

arXiv:2604.17969v1 Announce Type: new Abstract: Visual search in 3D environments requires embodied agents to actively explore their surroundings and acquire task-relevant evidence. However, existing v

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

EasyVideoR1: Easier RL for Video Understanding

DGX agent

arXiv:2604.16893v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) has demonstrated remarkable effectiveness in improving the reasoning capabilities of large languag

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

EchoChain: A Full-Duplex Benchmark for State-Update Reasoning Under Interruptions

DGX agent

arXiv:2604.16456v1 Announce Type: new Abstract: Real-time voice assistants must revise task state when users interrupt mid-response, but existing spoken-dialog benchmarks largely evaluate turn-based i

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning

DGX agent

arXiv:2509.20360v3 Announce Type: replace Abstract: Recent advances in foundation models highlight a clear trend toward unification and scaling, showing emergent capabilities across diverse domains. W

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Efficient Task Adaptation in Large Language Models via Selective Parameter Optimization

DGX agent

arXiv:2604.17051v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated excellent performance in general language understanding, generation and other tasks. However, when fine-t

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

EgoSound: Benchmarking Sound Understanding in Egocentric Videos

DGX agent

arXiv:2602.14122v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have recently achieved remarkable progress in vision-language understanding. Yet, human perception is inher

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Embedding Arithmetic: A Lightweight, Tuning-Free Framework for Post-hoc Bias Mitigation in Text-to-Image Models

DGX agent

arXiv:2604.18167v1 Announce Type: new Abstract: Modern text-to-image (T2I) models amplify harmful societal biases, challenging their ethical deployment. We introduce an inference-time method that reli

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents

DGX agent

arXiv:2604.18271v1 Announce Type: new Abstract: As the world of agentic artificial intelligence applied to robotics evolves, the need for agents capable of building and retrieving memories and observa

model-releasesarxiv-cs-ro
21 Apr 2026
Model Releases

Emergent Misalignment via In-Context Learning: Narrow in-context examples can produce broadly misaligned LLMs

DGX agent

arXiv:2510.11288v4 Announce Type: replace Abstract: Recent work has shown that narrow finetuning can produce broadly misaligned LLMs, a phenomenon termed emergent misalignment (EM). While concerning,

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Employing General-Purpose and Biomedical Large Language Models with Advanced Prompt Engineering for Pharmacoepidemiologic Study Design

DGX agent

arXiv:2604.17988v1 Announce Type: new Abstract: Background: The potential of large language models (LLMs) to automate and support pharmacoepidemiologic study design is an emerging area of interest, ye

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

End-to-end Listen, Look, Speak and Act

DGX agent

arXiv:2510.16756v2 Announce Type: replace-cross Abstract: Human interaction is inherently multimodal and full-duplex: we listen while watching, speak while acting, and fluidly adapt to turn-taking and

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Enhancing Continual Learning of Vision-Language Models via Dynamic Prefix Weighting

DGX agent

arXiv:2604.18075v1 Announce Type: new Abstract: We investigate recently introduced domain-class incremental learning scenarios for vision-language models (VLMs). Recent works address this challenge us

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Enhancing Glass Surface Reconstruction via Depth Prior for Robot Navigation

DGX agent

arXiv:2604.18336v1 Announce Type: cross Abstract: Indoor robot navigation is often compromised by glass surfaces, which severely corrupt depth sensor measurements. While foundation models like Depth A

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

ENTIRE: Learning-based Volume Rendering Time Prediction

DGX agent

arXiv:2501.12119v3 Announce Type: replace-cross Abstract: We introduce ENTIRE, a novel deep learning-based approach for fast and accurate volume rendering time prediction. Predicting rendering time is

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Error as Signal: Stiffness-Aware Diffusion Sampling via Embedded Runge-Kutta Guidance

DGX agent

arXiv:2603.03692v2 Announce Type: replace Abstract: Classifier-Free Guidance (CFG) has established the foundation for guidance mechanisms in diffusion models, showing that well-designed guidance proxi

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection

DGX agent

arXiv:2410.04509v3 Announce Type: replace Abstract: As the field of Multimodal Large Language Models (MLLMs) continues to evolve, their potential to revolutionize artificial intelligence is particular

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ESsEN: Training Compact Discriminative Vision-Language Transformers in a Low-Resource Setting

DGX agent

arXiv:2604.18452v1 Announce Type: cross Abstract: Vision-language modeling is rapidly increasing in popularity with an ever expanding list of available models. In most cases, these vision-language mod

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Evaluating Multimodal LLMs for Inpatient Diagnosis: Real-World Performance, Safety, and Cost Across Ten Frontier Models

DGX agent

arXiv:2604.16980v1 Announce Type: new Abstract: Background: Large language models (LLMs) are increasingly proposed for diagnostic support, but few evaluations use real-world multimodal inpatient data,

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Evaluating Tool-Using Language Agents: Judge Reliability, Propagation Cascades, and Runtime Mitigation in AgentProp-Bench

DGX agent

arXiv:2604.16706v1 Announce Type: cross Abstract: Automated evaluation of tool-using large language model (LLM) agents is widely assumed to be reliable, but this assumption has rarely been validated a

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

EvoCoT: Overcoming the Exploration Bottleneck in Reinforcement Learning

DGX agent

arXiv:2508.07809v5 Announce Type: replace Abstract: Reinforcement learning with verifiable reward (RLVR) has become a promising paradigm for post-training large language models (LLMs) to improve their

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Exploring Boundary-Aware Spatial-Frequency Fusion for Camouflaged Object Detection

DGX agent

arXiv:2604.17879v1 Announce Type: new Abstract: Camouflaged Object Detection is challenging due to the high degree of similarity between camouflaged objects and their surrounding backgrounds. Current

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

FaithLens: Detecting and Explaining Faithfulness Hallucination

DGX agent

arXiv:2512.20182v3 Announce Type: replace Abstract: Recognizing whether outputs from large language models (LLMs) contain faithfulness hallucination is crucial for real-world applications, e.g., retri

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Falcon flies every few days. You can see launches in person from Florida or California.

DGX agent

SpaceX's Falcon rockets conduct frequent launches occurring every few days, with public viewing opportunities available at launch facilities in Florida and California. This statement reflects SpaceX's

model-releaseselon-musk--x
21 Apr 2026
Model Releases

FedLLM: A Privacy-Preserving Federated Large Language Model for Explainable Traffic Flow Prediction

DGX agent

arXiv:2604.16612v1 Announce Type: new Abstract: Traffic prediction plays a central role in intelligent transportation systems (ITS) by supporting real-time decision-making, congestion management, and

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

FedOBP: Federated Optimal Brain Personalization through Cloud-Edge Element-wise Decoupling

DGX agent

arXiv:2604.16574v1 Announce Type: new Abstract: Federated Learning (FL) faces challenges from client data heterogeneity and resource-constrained mobile devices, which can degrade model accuracy. Perso

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Finding Culture-Sensitive Neurons in Vision-Language Models

DGX agent

arXiv:2510.24942v2 Announce Type: replace-cross Abstract: Despite their impressive performance, vision-language models (VLMs) still struggle on culturally situated inputs. To understand how VLMs proce

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

FireScope: Wildfire Risk Prediction with a Chain-of-Thought Oracle

DGX agent

arXiv:2511.17171v4 Announce Type: replace Abstract: Predicting wildfire risk is a reasoning-intensive spatial problem that requires the integration of visual, climatic, and geographic factors to infer

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

FLARE: A Data-Efficient Surrogate for Predicting Displacement Fields in Directed Energy Deposition

DGX agent

arXiv:2604.16649v1 Announce Type: new Abstract: Directed energy deposition (DED) produces complex thermo-mechanical responses that can lead to distortion and reduced dimensional accuracy of a manufact

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

FLARE: Task-agnostic embedding model evaluation through a normalization process

DGX agent

arXiv:2604.17344v1 Announce Type: cross Abstract: When task-specific labels are not available, it becomes difficult to select an embedding model for a specific target corpus. Existing labelless measur

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

FlashFPS: Efficient Farthest Point Sampling for Large-Scale Point Clouds via Pruning and Caching

DGX agent

arXiv:2604.17720v1 Announce Type: cross Abstract: Point-based Neural Networks (PNNs) have become a key approach for point cloud processing. However, a core operation in these models, Farthest Point Sa

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready t…

DGX agent

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready to fit the formats you need, from wide banners and presentati

model-releasesopenai--x
21 Apr 2026
Model Releases

FLiP: Towards understanding and interpreting multimodal multilingual sentence embeddings

DGX agent

arXiv:2604.18109v1 Announce Type: new Abstract: This paper presents factorized linear projection (FLiP) models for understanding pretrained sentence embedding spaces. We train FLiP models to recover t

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Flow marching for a generative PDE foundation model

DGX agent

arXiv:2509.18611v2 Announce Type: replace Abstract: Pretraining on large-scale collections of PDE-governed spatiotemporal trajectories has recently shown promise for building generalizable models of d

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

For clarity, we're running a small test on ~2% of new prosumer signups. Existing Pro and Max subscribers aren't affected.

DGX agent

For clarity, we're running a small test on ~2% of new prosumer signups. Existing Pro and Max subscribers aren't affected. Anthropic just pulled Claude Code from the Pro plan. Pro users wanting it need

model-releasesboris-cherny--x
21 Apr 2026
← Previous
1…407408409410411…466
Next →