AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,272 results
Model Releases

DERM-3R: A Resource-Efficient Multimodal Agents Framework for Dermatologic Diagnosis and Treatment in Real-World Clinical Settings

DGX agent

arXiv:2604.09596v1 Announce Type: new Abstract: Dermatologic diseases impose a large and growing global burden, affecting billions and substantially reducing quality of life. While modern therapies ca

model-releasesarxiv-cs-ai
14 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Design Principles for Sequence Models via Coefficient Dynamics

DGX agent

arXiv:2510.09389v2 Announce Type: replace-cross Abstract: Deep sequence models, ranging from Transformers and State Space Models (SSMs) to more recent approaches such as gated linear RNNs, fundamental

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Detecting Corporate AI-Washing via Cross-Modal Semantic Inconsistency Learning

DGX agent

arXiv:2604.09644v1 Announce Type: cross Abstract: Corporate AI-washing-the strategic misrepresentation of AI capabilities via exaggerated or fabricated cross-channel disclosures-has emerged as a syste

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Detecting critical treatment effect bias in small subgroups

DGX agent

arXiv:2404.18905v3 Announce Type: replace-cross Abstract: Randomized trials are considered the gold standard for making informed decisions in medicine, yet they often lack generalizability to the pati

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Detecting Safety Violations Across Many Agent Traces

DGX agent

arXiv:2604.11806v1 Announce Type: new Abstract: To identify safety violations, auditors often search over large sets of agent traces. This search is difficult because failures are often rare, complex,

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Development and evaluation of CADe systems in low-prevalence setting: The RARE25 challenge for early detection of Barrett's neoplasia

DGX agent

arXiv:2604.11171v1 Announce Type: new Abstract: Computer-aided detection (CADe) of early neoplasia in Barrett's esophagus is a low-prevalence surveillance problem in which clinically relevant findings

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Differentially Private Verification of Distribution Properties

DGX agent

arXiv:2604.10819v1 Announce Type: cross Abstract: A recent line of work initiated by Chiesa and Gur and further developed by Herman and Rothblum investigates the sample and communication complexity of

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

DiningBench: A Hierarchical Multi-view Benchmark for Perception and Reasoning in the Dietary Domain

DGX agent

arXiv:2604.10425v1 Announce Type: new Abstract: Recent advancements in Vision-Language Models (VLMs) have revolutionized general visual understanding. However, their application in the food domain rem

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Disambiguation-Centric Finetuning Makes Enterprise Tool-Calling LLMs More Realistic and Less Risky

DGX agent

arXiv:2507.03336v4 Announce Type: replace Abstract: Large language models (LLMs) are increasingly tasked with invoking enterprise APIs, yet they routinely falter when near-duplicate tools vie for the

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

DiSPA: Differential Substructure-Pathway Attention for Drug Response Prediction

DGX agent

arXiv:2601.14346v2 Announce Type: replace-cross Abstract: Accurate prediction of drug response in precision medicine requires models that capture how specific chemical substructures interact with cell

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Do Agent Rules Shape or Distort? Guardrails Beat Guidance in Coding Agents

DGX agent

arXiv:2604.11088v1 Announce Type: new Abstract: Developers increasingly guide AI coding agents through natural language instruction files (e.g., CLAUDE.md, .cursorrules), yet no controlled study has m

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Do LLMs Build Spatial World Models? Evidence from Grid-World Maze Tasks

DGX agent

arXiv:2604.10690v1 Announce Type: new Abstract: Foundation models have shown remarkable performance across diverse tasks, yet their ability to construct internal spatial world models for reasoning and

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Do Thought Streams Matter? Evaluating Reasoning in Gemini Vision-Language Models for Video Scene Understanding

DGX agent

arXiv:2604.11177v1 Announce Type: new Abstract: We benchmark how internal reasoning traces, which we call thought streams, affect video scene understanding in vision-language models. Using four config

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Do We Still Need GraphRAG? Benchmarking RAG and GraphRAG for Agentic Search Systems

DGX agent

arXiv:2604.09666v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) and its graph-based extensions (GraphRAG) are effective paradigms for improving large language model (LLM) reason

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models

DGX agent

arXiv:2601.03926v2 Announce Type: replace Abstract: The deployment of Large Vision-Language Models (LVLMs) for real-world document question answering is often constrained by dynamic, user-defined poli

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

DocRevive: A Unified Pipeline for Document Text Restoration

DGX agent

arXiv:2604.10077v1 Announce Type: new Abstract: In Document Understanding, the challenge of reconstructing damaged, occluded, or incomplete text remains a critical yet unexplored problem. Subsequent d

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Domain-Aware Hybrid Quantum Learning via Correlation-Guided Circuit Design for Crime Pattern Analytics

DGX agent

arXiv:2604.07389v2 Announce Type: replace Abstract: Crime pattern analysis is critical for law enforcement and predictive policing, yet the surge in criminal activities from rapid urbanization creates

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

DoReMi: Bridging 3D Domains via Topology-Aware Domain-Representation Mixture of Experts

DGX agent

arXiv:2511.11232v2 Announce Type: replace Abstract: Constructing a unified 3D scene understanding model has long been hindered by the significant topological discrepancies across different sensor moda

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

DPNet: Doppler LiDAR Motion Planning for Highly-Dynamic Environments

DGX agent

arXiv:2512.00375v2 Announce Type: replace Abstract: Existing motion planning methods often struggle with rapid-motion obstacles due to an insufficient understanding of environmental changes. To addres

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

E2E-REME: Towards End-to-End Microservices Auto-Remediation via Experience-Simulation Reinforcement Fine-Tuning

DGX agent

arXiv:2604.11094v1 Announce Type: cross Abstract: Contemporary microservice systems continue to grow in scale and complexity, leading to increasingly frequent and costly failures. While recent LLM-bas

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

EagleVision: A Multi-Task Benchmark for Cross-Domain Perception in High-Speed Autonomous Racing

DGX agent

arXiv:2604.11400v1 Announce Type: cross Abstract: High-speed autonomous racing presents extreme perception challenges, including large relative velocities and substantial domain shifts from convention

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

EdgeCIM: A Hardware-Software Co-Design for CIM-Based Acceleration of Small Language Models

DGX agent

arXiv:2604.11512v1 Announce Type: cross Abstract: The growing demand for deploying Small Language Models (SLMs) on edge devices, including laptops, smartphones, and embedded platforms, has exposed fun

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

EdgeDAM: Real-time Object Tracking for Mobile Devices

DGX agent

arXiv:2603.05463v2 Announce Type: replace Abstract: Single-object tracking (SOT) on edge devices is a critical computer vision task, requiring accurate and continuous target localization across video

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Edu-MMBias: A Three-Tier Multimodal Benchmark for Auditing Social Bias in Vision-Language Models under Educational Contexts

DGX agent

arXiv:2604.10200v1 Announce Type: new Abstract: As Vision-Language Models (VLMs) become integral to educational decision-making, ensuring their fairness is paramount. However, current text-centric eva

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

EduIllustrate: Towards Scalable Automated Generation Of Multimodal Educational Content

DGX agent

arXiv:2604.05005v2 Announce Type: replace-cross Abstract: Large language models are increasingly used as educational assistants, yet evaluation of their educational capabilities remains concentrated o

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Eevee: Towards Close-up High-resolution Video-based Virtual Try-on

DGX agent

arXiv:2511.18957v2 Announce Type: replace Abstract: Video virtual try-on technology provides a cost-effective solution for creating marketing videos in fashion e-commerce. However, its practical adopt

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

EgoFun3D: Modeling Interactive Objects from Egocentric Videos using Function Templates

DGX agent

arXiv:2604.11038v1 Announce Type: new Abstract: We present EgoFun3D, a coordinated task formulation, dataset, and benchmark for modeling interactive 3D objects from egocentric videos. Interactive obje

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Eliciting Medical Reasoning with Knowledge-enhanced Data Synthesis: A Semi-Supervised Reinforcement Learning Approach

DGX agent

arXiv:2604.11547v1 Announce Type: cross Abstract: While large language models hold promise for complex medical applications, their development is hindered by the scarcity of high-quality reasoning dat

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems

DGX agent

arXiv:2604.11174v1 Announce Type: cross Abstract: Recent progress in embodied AI has produced a growing ecosystem of robot policies, foundation models, and modular runtimes. However, current evaluatio

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

End-to-end Automated Deep Neural Network Optimization for PPG-based Blood Pressure Estimation on Wearables

DGX agent

arXiv:2604.10117v1 Announce Type: new Abstract: Photoplethysmography (PPG)-based blood pressure (BP) estimation is a challenging task, particularly on resource-constrained wearable devices. However, f

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Enhanced-FQL(lambda), an Efficient and Interpretable RL with novel Fuzzy Eligibility Traces and Segmented Experience Replay

DGX agent

arXiv:2601.04392v2 Announce Type: replace-cross Abstract: This paper introduces a fuzzy reinforcement learning framework, Enhanced-FQL(lambda), that integrates novel Fuzzified Eligibility Traces (FET)

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Enhancing Multimodal Large Language Models for Ancient Chinese Character Evolution Analysis via Glyph-Driven Fine-Tuning

DGX agent

arXiv:2604.11299v1 Announce Type: cross Abstract: In recent years, rapid advances in Multimodal Large Language Models (MLLMs) have increasingly stimulated research on ancient Chinese scripts. As the e

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Environmental Footprint of GenAI Research: Insights from the Moshi Foundation Model

DGX agent

arXiv:2604.11154v1 Announce Type: new Abstract: New multi-modal large language models (MLLMs) are continuously being trained and deployed, following rapid development cycles. This generative AI frenzy

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ERNIE Image released

DGX agent

ERNIE Image is an open-source text-to-image generation model developed by Baidu, built on a single-stream Diffusion Transformer (DiT) paired with a lightweight Prompt Enhancer that expands brief user

model-releasesr-stablediffusion
14 Apr 2026
Model Releases

Escaping the Context Bottleneck: Active Context Curation for LLM Agents via Reinforcement Learning

DGX agent

arXiv:2604.11462v1 Announce Type: new Abstract: Large Language Models (LLMs) struggle with long-horizon tasks due to the 'context bottleneck' and the 'lost-in-the-middle' phenomenon, where accumulated

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Evaluating Memory Capability in Continuous Lifelog Scenario

DGX agent

arXiv:2604.11182v1 Announce Type: new Abstract: Nowadays, wearable devices can continuously lifelog ambient conversations, creating substantial opportunities for memory systems. However, existing benc

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Evaluating Reliability Gaps in Large Language Model Safety via Repeated Prompt Sampling

DGX agent

arXiv:2604.09606v1 Announce Type: new Abstract: Traditional benchmarks for large language models (LLMs), such as HELM and AIR-BENCH, primarily assess safety risk through breadth-oriented evaluation ac

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Evaluating Scene-based In-Situ Item Labeling for Immersive Conversational Recommendation

DGX agent

arXiv:2604.09698v1 Announce Type: cross Abstract: The growing ubiquity of Extended Reality (XR) is driving Conversational Recommendation Systems (CRS) toward visually immersive experiences. We formali

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Evaluating Small Open LLMs for Medical Question Answering: A Practical Framework

DGX agent

arXiv:2604.10535v1 Announce Type: cross Abstract: Incorporating large language models (LLMs) in medical question answering demands more than high average accuracy: a model that returns substantively d

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

EviRCOD: Evidence-Guided Probabilistic Decoding for Referring Camouflaged Object Detection

DGX agent

arXiv:2604.10894v1 Announce Type: new Abstract: Referring Camouflaged Object Detection (Ref-COD) focuses on segmenting specific camouflaged targets in a query image using category-aligned references.

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

EvoDiagram: Agentic Editable Diagram Creation via Design Expertise Evolution

DGX agent

arXiv:2604.09568v1 Announce Type: cross Abstract: High-fidelity diagram creation requires the complex orchestration of semantic topology, visual styling, and spatial layout, posing a significant chall

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Excited to be part of @ollama Gemma Day tomorrow in Palo Alto! Khoa Pham (@kwafam7) from @radixark will demo Gemma 4 in production with SGLa…

DGX agent

Excited to be part of @ollama Gemma Day tomorrow in Palo Alto! Khoa Pham (@kwafam7) from @radixark will demo Gemma 4 in production with SGLang 🚀 Come say hi! 📷 RSVP: https://luma.com/ollama-gemma4 Oll

model-releasesollama--x
14 Apr 2026
Model Releases

ExecTune: Effective Steering of Black-Box LLMs with Guide Models

DGX agent

arXiv:2604.09741v1 Announce Type: cross Abstract: For large language models deployed through black-box APIs, recurring inference costs often exceed one-time training costs. This motivates composed age

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Exploring Cross-Modal Flows for Few-Shot Learning

DGX agent

arXiv:2510.14543v4 Announce Type: replace Abstract: Aligning features from different modalities, is one of the most fundamental challenges for cross-modal tasks. Although pre-trained vision-language m

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Exploring Knowledge Conflicts for Faithful LLM Reasoning: Benchmark and Method

DGX agent

arXiv:2604.11209v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable success across a wide range of applications especially when augmented by external knowledge thro

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Exploring the best way for UAV visual localization under Low-altitude Multi-view Observation Condition: a Benchmark

DGX agent

arXiv:2503.10692v2 Announce Type: replace Abstract: Absolute Visual Localization (AVL) enables an Unmanned Aerial Vehicle (UAV) to determine its position in GNSS-denied environments by establishing ge

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Falcon 9 launches 29 @Starlink satellites from Florida

DGX agent

SpaceX's Falcon 9 rocket successfully launched 29 Starlink satellites from a Florida launch site, further expanding SpaceX's broadband internet constellation. This mission is part of SpaceX's ongoing

model-releaseselon-musk--x
14 Apr 2026
Model Releases

FashionMV: Product-Level Composed Image Retrieval with Multi-View Fashion Data

DGX agent

arXiv:2604.10297v1 Announce Type: cross Abstract: Composed Image Retrieval (CIR) retrieves target images using a reference image paired with modification text. Despite rapid advances, all existing met

model-releasesarxiv-cs-ai
14 Apr 2026
← Previous
1…437438439440441…464
Next →