AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,098 results
10 Apr 2026

Commander-GPT: Dividing and Routing for Multimodal Sarcasm Detection

Model ReleasesDGX agent

arXiv:2506.19420v2 Announce Type: replace Abstract: Multimodal sarcasm understanding is a high-order cognitive task. Although large language models (LLMs) have shown impressive performance on many dow

Consistency-Guided Decoding with Proof-Driven Disambiguation for Three-Way Logical Question Answering

Model ReleasesDGX agent

arXiv:2604.06196v1 Announce Type: cross Abstract: Three-way logical question answering (QA) assigns True/False/Unknown to a hypothesis H given a premise set S. While modern large language models

ConsistRM: Improving Generative Reward Models via Consistency-Aware Self-Training

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.07484v1 Announce Type: cross Abstract: Generative reward models (GRMs) have emerged as a promising approach for aligning Large Language Models (LLMs) with human preferences by offering grea

Contextual Earnings-22: A Speech Recognition Benchmark with Custom Vocabulary in the Wild

Model ReleasesDGX agent

arXiv:2604.07354v1 Announce Type: new Abstract: The accuracy frontier of speech-to-text systems has plateaued on academic benchmarks.1 In contrast, industrial benchmarks and adoption in high-stakes do

Continual Visual Anomaly Detection on the Edge: Benchmark and Efficient Solutions

Model ReleasesDGX agent

arXiv:2604.06435v1 Announce Type: cross Abstract: Visual Anomaly Detection (VAD) is a critical task for many applications including industrial inspection and healthcare. While VAD has been extensively

ConvoLearn: A Dataset for Fine-Tuning Dialogic AI Tutors

Model ReleasesDGX agent

arXiv:2601.08950v3 Announce Type: replace Abstract: Despite their growing adoption in education, LLMs remain misaligned with the core principle of effective tutoring: the dialogic construction of know

CoreWeave inks multiyear cloud deal with Anthropic

Model ReleasesDGX agent

CoreWeave Inc. today announced that it has won a multiyear contract to supply Anthropic PBC with cloud infrastructure. The company’s shares closed 11% higher on the news. The data center capacity comm

CrashSight: A Phase-Aware, Infrastructure-Centric Video Benchmark for Traffic Crash Scene Understanding and Reasoning

Model ReleasesDGX agent

arXiv:2604.08457v1 Announce Type: new Abstract: Cooperative autonomous driving requires traffic scene understanding from both vehicle and infrastructure perspectives. While vision-language models (VLM

Create Expert Content: Local Testing of a Multi-Agent System with Memory

Model ReleasesDGX agent

In support of our mission to accelerate the developer journey on Google Cloud, we built Dev Signal: a multi-agent system designed to transform raw community signals into reliable technical guidance by

Cron Jobs — AI That Works on a Schedule Tell Qwen Code 'check if tests pass every 30 minutes' and it sets up a cron job in your session. No …

Model ReleasesDGX agent

Cron Jobs — AI That Works on a Schedule Tell Qwen Code 'check if tests pass every 30 minutes' and it sets up a cron job in your session. No crontab editing, no scripts to write. Use /loop commend. Wor

Cross-Lingual Transfer and Parameter-Efficient Adaptation in the Turkic Language Family: A Theoretical Framework for Low-Resource Language Models

Model ReleasesDGX agent

arXiv:2604.06202v1 Announce Type: cross Abstract: Large language models (LLMs) have transformed natural language processing, yet their capabilities remain uneven across languages. Most multilingual mo

CryoSplat: Gaussian Splatting for Cryo-EM Homogeneous Reconstruction

Model ReleasesDGX agent

arXiv:2508.04929v4 Announce Type: replace-cross Abstract: As a critical modality for structural biology, cryogenic electron microscopy (cryo-EM) facilitates the determination of macromolecular structu

CycleChart: A Unified Consistency-Based Learning Framework for Bidirectional Chart Understanding and Generation

Model ReleasesDGX agent

arXiv:2512.19173v2 Announce Type: replace Abstract: Current chart-related tasks, such as chart generation (NL2Chart), chart schema parsing, chart data parsing, and chart question answering (ChartQA),

DDP-SA: Scalable Privacy-Preserving Federated Learning via Distributed Differential Privacy and Secure Aggregation

Model ReleasesDGX agent

arXiv:2604.07125v1 Announce Type: cross Abstract: This article presents DDP-SA, a scalable privacy-preserving federated learning framework that jointly leverages client-side local differential privacy

Dear MAGA, Just out of curiosity: If president Biden had banged porn stars, cheated on multiple wives, lied 30,000 times, singlehandedly and…

Model ReleasesDGX agent

Dear MAGA, Just out of curiosity: If president Biden had banged porn stars, cheated on multiple wives, lied 30,000 times, singlehandedly and unilaterally started a war, released 5,000 Talibanis, bombe

Deep Agents Deploy: an open alternative to Claude Managed Agents

Model ReleasesDGX agent

LangChain launched **Deep Agents Deploy** in beta as an open-source, model-agnostic alternative to Anthropic's Claude Managed Agents. It is designed to be the fastest way to deploy a model-agnosti...

Detecting HIV-Related Stigma in Clinical Narratives Using Large Language Models

Model ReleasesDGX agent

arXiv:2604.07717v1 Announce Type: new Abstract: Human immunodeficiency virus (HIV)-related stigma is a critical psychosocial determinant of health for people living with HIV (PLWH), influencing mental

Diagnosing and Mitigating Sycophancy and Skepticism in LLM Causal Judgment

Model ReleasesDGX agent

arXiv:2601.08258v3 Announce Type: replace Abstract: Large language models increasingly fail in a way that scalar accuracy cannot diagnose: they produce a sound reasoning trace and then abandon it unde

Digital Skin, Digital Bias: Uncovering Tone-Based Biases in LLMs and Emoji Embeddings

Model ReleasesDGX agent

arXiv:2604.06863v1 Announce Type: cross Abstract: Skin-toned emojis are crucial for fostering personal identity and social inclusion in online communication. As AI models, particularly Large Language

Direct Segmentation without Logits Optimization for Training-Free Open-Vocabulary Semantic Segmentation

Model ReleasesDGX agent

arXiv:2604.07723v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation (OVSS) aims to segment arbitrary category regions in images using open-vocabulary prompts, necessitating that exis

DISSECT: Diagnosing Where Vision Ends and Language Priors Begin in Scientific VLMs

Model ReleasesDGX agent

arXiv:2604.06250v1 Announce Type: cross Abstract: When asked to describe a molecular diagram, a Vision-Language Model correctly identifies ``a benzene ring with an -OH group.'' When asked to reason ab

Distributed Interpretability and Control for Large Language Models

Model ReleasesDGX agent

arXiv:2604.06483v1 Announce Type: cross Abstract: Large language models that require multiple GPU cards to host are usually the most capable models. It is necessary to understand and steer these model

Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models

Model ReleasesDGX agent

arXiv:2604.08284v1 Announce Type: new Abstract: Large language models store not only isolated facts but also rules that support reasoning across symbolic expressions, natural language explanations, an

Do MLLMs Really Understand Space? A Mathematical Reasoning Evaluation

Model ReleasesDGX agent

arXiv:2602.11635v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have achieved strong performance on perception-oriented tasks, yet their ability to perform mathematical sp

Domain-Contextualized Inference: A Computable Graph Architecture for Explicit-Domain Reasoning

Model ReleasesDGX agent

arXiv:2604.04344v2 Announce Type: replace Abstract: We establish a computation-substrate-agnostic inference architecture in which domain is an explicit first-class computational parameter. This produc

Don't Overthink It: Inter-Rollout Action Agreement as a Free Adaptive-Compute Signal for LLM Agents

Model ReleasesDGX agent

arXiv:2604.08369v1 Announce Type: cross Abstract: Inference-time compute scaling has emerged as a powerful technique for improving the reliability of large language model (LLM) agents, but existing me

DosimeTron: Automating Personalized Monte Carlo Radiation Dosimetry in PET/CT with Agentic AI

Model ReleasesDGX agent

arXiv:2604.06280v1 Announce Type: cross Abstract: Purpose: To develop and evaluate DosimeTron, an agentic AI system for automated patient-specific MC internal radiation dosimetry in PET/CT examination

Draw-In-Mind: Rebalancing Designer-Painter Roles in Unified Multimodal Models Benefits Image Editing

Model ReleasesDGX agent

arXiv:2509.01986v4 Announce Type: replace-cross Abstract: In recent years, integrating multimodal understanding and generation into a single unified model has emerged as a promising paradigm. While th

DROP: Distributional and Regular Optimism and Pessimism for Reinforcement Learning

Model ReleasesDGX agent

arXiv:2410.17473v2 Announce Type: replace Abstract: In reinforcement learning (RL), temporal difference (TD) error is known to be related to the firing rate of dopamine neurons. It has been observed t

DSCA: Dynamic Subspace Concept Alignment for Lifelong VLM Editing

Model ReleasesDGX agent

arXiv:2604.07965v1 Announce Type: new Abstract: Model editing aims to update knowledge to add new concepts and change relevant information without retraining. Lifelong editing is a challenging task, p

Dual-level Modality Debiasing Learning for Unsupervised Visible-Infrared Person Re-Identification

Model ReleasesDGX agent

arXiv:2512.03745v2 Announce Type: replace Abstract: Two-stage learning pipeline has achieved promising results in unsupervised visible-infrared person re-identification (USL-VI-ReID). It first perform

Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Serving

Model ReleasesDGX agent

arXiv:2604.08075v1 Announce Type: new Abstract: Production vLLM fleets typically provision each instance for the worst-case context length, leading to substantial KV-cache over-allocation and under-ut

Dynamic Context Evolution for Scalable Synthetic Data Generation

Model ReleasesDGX agent

arXiv:2604.07147v1 Announce Type: cross Abstract: Large language models produce repetitive output when prompted independently across many batches, a phenomenon we term cross-batch mode collapse: the p

E2Edev: Benchmarking Large Language Models in End-to-End Software Development Task

Model ReleasesDGX agent

arXiv:2510.14509v3 Announce Type: replace-cross Abstract: The rapid advancement in large language models (LLMs) has demonstrated significant potential in End-to-End Software Development (E2ESD). Howev

EditCaption: Human-Aligned Instruction Synthesis for Image Editing via Supervised Fine-Tuning and Direct Preference Optimization

Model ReleasesDGX agent

arXiv:2604.08213v1 Announce Type: new Abstract: High-quality training triplets (source-target image pairs with precise editing instructions) are a critical bottleneck for scaling instruction-guided im

Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Prediction

Model ReleasesDGX agent

arXiv:2604.07659v1 Announce Type: new Abstract: Large language models (LLMs) hold significant promise for healthcare, yet their reliability in high-stakes clinical settings is often compromised by hal

Emotion Concepts and their Function in a Large Language Model

Model ReleasesDGX agent

arXiv:2604.07729v1 Announce Type: cross Abstract: Large language models (LLMs) sometimes appear to exhibit emotional reactions. We investigate why this is the case in Claude Sonnet 4.5 and explore imp

Enabling Intrinsic Reasoning over Dense Geospatial Embeddings with DFR-Gemma

Model ReleasesDGX agent

arXiv:2604.07490v1 Announce Type: new Abstract: Representation learning for geospatial and spatio-temporal data plays a critical role in enabling general-purpose geospatial intelligence. Recent geospa

Entropy After </Think> for reasoning model early exiting

Model ReleasesDGX agent

arXiv:2509.26522v3 Announce Type: replace Abstract: Reasoning LLMs show improved performance with longer chains of thought. However, recent work has highlighted their tendency to overthink, continuing

Epistemic Robust Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.07072v1 Announce Type: new Abstract: Offline reinforcement learning learns policies from fixed datasets without further environment interaction. A key challenge in this setting is epistemic

ESOM: Efficiently Understanding Streaming Video Anomalies with Open-world Dynamic Definitions

Model ReleasesDGX agent

arXiv:2604.07772v1 Announce Type: new Abstract: Open-world video anomaly detection (OWVAD) aims to detect and explain abnormal events under different anomaly definitions, which is important for applic

ETCH-X: Robustify Expressive Body Fitting to Clothed Humans with Composable Datasets

Model ReleasesDGX agent

arXiv:2604.08548v1 Announce Type: new Abstract: Human body fitting, which aligns parametric body models such as SMPL to raw 3D point clouds of clothed humans, serves as a crucial first step for downst

Evaluating LLM-Based 0-to-1 Software Generation in End-to-End CLI Tool Scenarios

Model ReleasesDGX agent

arXiv:2604.06742v1 Announce Type: cross Abstract: Large Language Models (LLMs) are driving a shift towards intent-driven development, where agents build complete software from scratch. However, existi

Evaluating LLMs for Demographic-Targeted Social Bias Detection: A Comprehensive Benchmark Study

Model ReleasesDGX agent

arXiv:2510.04641v3 Announce Type: replace Abstract: Large-scale web-scraped text corpora used to train general-purpose AI models often contain harmful demographic-targeted social biases, creating a re

Evaluating Low-Light Image Enhancement Across Multiple Intensity Levels

Model ReleasesDGX agent

arXiv:2511.15496v2 Announce Type: replace Abstract: Imaging in low-light environments is challenging due to reduced scene radiance, which leads to elevated sensor noise and reduced color saturation. M

Evaluating Repository-level Software Documentation via Question Answering and Feature-Driven Development

Model ReleasesDGX agent

arXiv:2604.06793v1 Announce Type: cross Abstract: Software documentation is crucial for repository comprehension. While Large Language Models (LLMs) advance documentation generation from code snippets

EVGeoQA: Benchmarking LLMs on Dynamic, Multi-Objective Geo-Spatial Exploration

Model ReleasesDGX agent

arXiv:2604.07070v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, their potential for purpose-driven exploration in dynamic geo-spatial

EvoGymCM: Harnessing Continuous Material Stiffness for Soft Robot Co-Design

Model ReleasesDGX agent

arXiv:2604.08258v1 Announce Type: new Abstract: In the automated co-design of soft robots, precisely adapting the material stiffness field to task environments is crucial for unlocking their full phys

Face-D(^2)CL: Multi-Domain Synergistic Representation with Dual Continual Learning for Facial DeepFake Detection

Model ReleasesDGX agent

arXiv:2604.08159v1 Announce Type: new Abstract: The rapid advancement of facial forgery techniques poses severe threats to public trust and information security, making facial DeepFake detection a cri

Fail2Drive: Benchmarking Closed-Loop Driving Generalization

Model ReleasesDGX agent

arXiv:2604.08535v1 Announce Type: cross Abstract: Generalization under distribution shift remains a central bottleneck for closed-loop autonomous driving. Although simulators like CARLA enable safe an

FedSpy-LLM: Towards Scalable and Generalizable Data Reconstruction Attacks from Gradients on LLMs

Model ReleasesDGX agent

arXiv:2604.06297v1 Announce Type: cross Abstract: Given the growing reliance on private data in training Large Language Models (LLMs), Federated Learning (FL) combined with Parameter-Efficient Fine-Tu

Fighting AI with AI: AI-Agent Augmented DNS Blocking of LLM Services during Student Evaluations

Model ReleasesDGX agent

arXiv:2604.02360v1 Announce Type: cross Abstract: The transformative potential of large language models (LLMs) in education, such as improving accessibility and personalized learning, is being eclipse

FinTruthQA: A Benchmark for AI-Driven Financial Disclosure Quality Assessment in Investor -- Firm Interactions

Model ReleasesDGX agent

arXiv:2406.12009v5 Announce Type: replace Abstract: Accurate and transparent financial information disclosure is essential for market efficiency, investor decision-making, and corporate governance. Ch

FireSenseNet: A Dual-Branch CNN with Cross-Attentive Feature Interaction for Next-Day Wildfire Spread Prediction

Model ReleasesDGX agent

arXiv:2604.07675v1 Announce Type: new Abstract: Accurate prediction of next-day wildfire spread is critical for disaster response and resource allocation. Existing deep learning approaches typically c

FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On

Model ReleasesDGX agent

arXiv:2604.08526v1 Announce Type: new Abstract: Given a person and a garment image, virtual try-on (VTO) aims to synthesize a realistic image of the person wearing the garment, while preserving their

FLeX: Fourier-based Low-rank EXpansion for multilingual transfer

Model ReleasesDGX agent

arXiv:2604.06253v1 Announce Type: cross Abstract: Cross-lingual code generation is critical in enterprise environments where multiple programming languages coexist. However, fine-tuning large language

Flow Motion Policy: Manipulator Motion Planning with Flow Matching Models

Model ReleasesDGX agent

arXiv:2604.07084v1 Announce Type: cross Abstract: Open-loop end-to-end neural motion planners have recently been proposed to improve motion planning for robotic manipulators. These methods enable plan

FlowAdam: Implicit Regularization via Geometry-Aware Soft Momentum Injection

Model ReleasesDGX agent

arXiv:2604.06652v1 Announce Type: new Abstract: Adaptive moment methods such as Adam use a diagonal, coordinate-wise preconditioner based on exponential moving averages of squared gradients. This diag

FlowGuard: Towards Lightweight In-Generation Safety Detection for Diffusion Models via Linear Latent Decoding

Model ReleasesDGX agent

arXiv:2604.07879v1 Announce Type: new Abstract: Diffusion-based image generation models have advanced rapidly but pose a safety risk due to their potential to generate Not-Safe-For-Work (NSFW) content

Flux Attention: Context-Aware Hybrid Attention for Efficient LLMs Inference

Model ReleasesDGX agent

arXiv:2604.07394v1 Announce Type: cross Abstract: The quadratic computational complexity of standard attention mechanisms presents a severe scalability bottleneck for LLMs in long-context scenarios. W

← Previous
1…360361362363364…369
Next →