AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

WorldMark: A Unified Benchmark Suite for Interactive Video World Models

DGX agent

arXiv:2604.21686v1 Announce Type: new Abstract: Interactive video generation models such as Genie, YUME, HY-World, and Matrix-Game are advancing rapidly, yet every model is evaluated on its own benchm

model-releasesarxiv-cs-cv
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

You Only Gaussian Once: Controllable 3D Gaussian Splatting for Ultra-Densely Sampled Scenes

DGX agent

arXiv:2604.21400v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has revolutionized neural rendering, yet existing methods remain predominantly research prototypes ill-suited for productio

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Zero-Shot Detection of LLM-Generated Text via Implicit Reward Model

DGX agent

arXiv:2604.21223v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated remarkable capabilities across various tasks. However, their ability to generate human-like text has ra

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

A Vision-Language-Action Model for Adaptive Ultrasound-Guided Needle Insertion and Needle Tracking

DGX agent

arXiv:2604.20347v1 Announce Type: cross Abstract: Ultrasound (US)-guided needle insertion is a critical yet challenging procedure due to dynamic imaging conditions and difficulties in needle visualiza

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

A weighted angle distance on strings

DGX agent

arXiv:2604.20633v1 Announce Type: cross Abstract: We define a multi-scale metric d_rho on strings by aggregating angle distances between all n-gram count vectors with exponential weights rho^n. We ben

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

AAC: Admissible-by-Architecture Differentiable Landmark Compression for ALT

DGX agent

arXiv:2604.20744v1 Announce Type: new Abstract: We introduce extbf{AAC} (Architecturally Admissible Compressor), a differentiable landmark-selection module for ALT (A*, Landmarks, and Triangle inequal

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Accelerating PayPal's Commerce Agent with Speculative Decoding: An Empirical Study on EAGLE3 with Fine-Tuned Nemotron Models

DGX agent

arXiv:2604.19767v1 Announce Type: cross Abstract: We evaluate speculative decoding with EAGLE3 as an inference-time optimization for PayPal's Commerce Agent, powered by a fine-tuned llama3.1-nemotron-

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

ActuBench: A Multi-Agent LLM Pipeline for Generation and Evaluation of Actuarial Reasoning Tasks

DGX agent

arXiv:2604.20273v1 Announce Type: new Abstract: We present ActuBench, a multi-agent LLM pipeline for the automated generation and evaluation of advanced actuarial assessment items aligned with the Int

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Adapting TrOCR for Printed Tigrinya Text Recognition: Word-Aware Loss Weighting for Cross-Script Transfer Learning

DGX agent

arXiv:2604.20813v1 Announce Type: new Abstract: Transformer-based OCR models have shown strong performance on Latin and CJK scripts, but their application to African syllabic writing systems remains l

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

AI to Learn 2.0: A Deliverable-Oriented Governance Framework and Maturity Rubric for Opaque AI in Learning-Intensive Domains

DGX agent

arXiv:2604.19751v1 Announce Type: new Abstract: Generative AI is entering research, education, and professional work faster than current governance frameworks can specify how AI-assisted outputs shoul

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Anchor-and-Resume Concession Under Dynamic Pricing for LLM-Augmented Freight Negotiation

DGX agent

arXiv:2604.20732v1 Announce Type: cross Abstract: Freight brokerages negotiate thousands of carrier rates daily under dynamic pricing conditions where models frequently revise targets mid-conversation

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Are LLM Uncertainty and Correctness Encoded by the Same Features? A Functional Dissociation via Sparse Autoencoders

DGX agent

arXiv:2604.19974v1 Announce Type: cross Abstract: Large language models can be uncertain yet correct, or confident yet wrong, raising the question of whether their output-level uncertainty and their a

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Assessing the Robustness of Climate Foundation Models under No-Analog Distribution Shifts

DGX agent

arXiv:2603.23043v2 Announce Type: replace-cross Abstract: The accelerating pace of climate change introduces profound non-stationarities that challenge the ability of Machine Learning based climate em

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

ATIR: Towards Audio-Text Interleaved Contextual Retrieval

DGX agent

arXiv:2604.20267v1 Announce Type: cross Abstract: Audio carries richer information than text, including emotion, speaker traits, and environmental context, while also enabling lower-latency processing

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Automated Detection of Dosing Errors in Clinical Trial Narratives: A Multi-Modal Feature Engineering Approach with LightGBM

DGX agent

arXiv:2604.19759v1 Announce Type: new Abstract: Clinical trials require strict adherence to medication protocols, yet dosing errors remain a persistent challenge affecting patient safety and trial int

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Automatic Ontology Construction Using LLMs as an External Layer of Memory, Verification, and Planning for Hybrid Intelligent Systems

DGX agent

arXiv:2604.20795v1 Announce Type: new Abstract: This paper presents a hybrid architecture for intelligent systems in which large language models (LLMs) are extended with an external ontological memory

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

AVISE: Framework for Evaluating the Security of AI Systems

DGX agent

arXiv:2604.20833v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across critical domains, their security vulnerabilities pose growing risks of high-p

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Benchmarking ResNet for Short-Term Hypoglycemia Classification with DiaData

DGX agent

arXiv:2511.02849v2 Announce Type: replace-cross Abstract: Individualized therapy is driven forward by medical data analysis, which provides insight into the patient's context. In particular, for Type

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Benefits of Low-Cost Bio-Inspiration in the Age of Overparametrization

DGX agent

arXiv:2604.20365v1 Announce Type: cross Abstract: While Central Pattern Generators (CPGs) and Multi-Layer Perceptrons (MLP) are widely used paradigms in robot control, few systematic studies have been

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Beyond Majority Voting: Towards Fine-grained and More Reliable Reward Signal for Test-Time Reinforcement Learning

DGX agent

arXiv:2512.15146v3 Announce Type: replace Abstract: Test-time reinforcement learning mitigates the reliance on annotated data by using majority voting results as pseudo-labels, emerging as a complemen

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models

DGX agent

arXiv:2604.16902v2 Announce Type: replace Abstract: Native Omni-modal Large Language Models (OLLMs) have shifted from pipeline architectures to unified representation spaces. However, this native inte

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Beyond the Crowd: LLM-Augmented Community Notes for Governing Health Misinformation

DGX agent

arXiv:2510.11423v3 Announce Type: replace-cross Abstract: Community Notes, the crowd-sourced misinformation governance system on X (formerly Twitter), allows users to flag misleading posts, attach con

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Bimanual Robot Manipulation via Multi-Agent In-Context Learning

DGX agent

arXiv:2604.20348v1 Announce Type: cross Abstract: Language Models (LLMs) have emerged as powerful reasoning engines for embodied control. In particular, In-Context Learning (ICL) enables off-the-shelf

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Bootstrapping Post-training Signals for Open-ended Tasks via Rubric-based Self-play on Pre-training Text

DGX agent

arXiv:2604.20051v1 Announce Type: new Abstract: Self-play has recently emerged as a promising paradigm to train Large Language Models (LLMs). In self-play, the target LLM creates the task input (e.g.,

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Bridging Mechanistic Interpretability and Prompt Engineering with Gradient Ascent for Interpretable Persona Control

DGX agent

arXiv:2601.02896v2 Announce Type: replace Abstract: Controlling emergent behavioral personas (e.g., sycophancy, hallucination) in Large Language Models (LLMs) is critical for AI safety, yet remains a

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

Can 'AI' Be a Doctor? A Study of Empathy, Readability, and Alignment in Clinical LLMs

DGX agent

arXiv:2604.20791v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed in healthcare, yet their communicative alignment with clinical standards remains insufficiently

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Can We Locate and Prevent Stereotypes in LLMs?

DGX agent

arXiv:2604.19764v1 Announce Type: cross Abstract: Stereotypes in large language models (LLMs) can perpetuate harmful societal biases. Despite the widespread use of models, little is known about where

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

CARLA-Air: Fly Drones Inside a CARLA World -- A Unified Infrastructure for Air-Ground Embodied Intelligence

DGX agent

arXiv:2603.28032v2 Announce Type: replace-cross Abstract: The convergence of low-altitude economies, embodied intelligence, and air-ground cooperative systems creates growing demand for simulation inf

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Catalyzing Informed Residential Energy Retrofit Decisions via Domain-Specific LLM

DGX agent

arXiv:2602.20181v2 Announce Type: replace-cross Abstract: Residential energy retrofit initiation is often stalled by an expertise gap, where homeowners lack the technical literacy required for structu

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

CCTVBench: Contrastive Consistency Traffic VideoQA Benchmark for Multimodal LLMs

DGX agent

arXiv:2604.20460v1 Announce Type: new Abstract: Safety-critical traffic reasoning requires contrastive consistency: models must detect true hazards when an accident occurs, and reliably reject plausib

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Chasing the Public Score: User Pressure and Evaluation Exploitation in Coding Agent Workflows

DGX agent

arXiv:2604.20200v1 Announce Type: new Abstract: Frontier coding agents are increasingly used in workflows where users supervise progress primarily through repeated improvement of a public score, namel

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

CLIP-SVD: Efficient and Interpretable Vision-Language Adaptation via Singular Values

DGX agent

arXiv:2509.03740v3 Announce Type: replace-cross Abstract: Vision-language models (VLMs) like CLIP have shown impressive zero-shot and few-shot learning capabilities across diverse applications. Howeve

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Coding with Eyes: Visual Feedback Unlocks Reliable GUI Code Generating and Debugging

DGX agent

arXiv:2604.19750v1 Announce Type: cross Abstract: Recent advances in Large Language Model (LLM)-based agents have shown remarkable progress in code generation. However, current agent methods mainly re

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Cognitive Kernel-Pro: A Framework for Deep Research Agents and Agent Foundation Models Training

DGX agent

arXiv:2508.00414v3 Announce Type: replace Abstract: General AI Agents are increasingly recognized as foundational frameworks for the next generation of artificial intelligence, enabling complex reason

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling

DGX agent

arXiv:2604.20720v1 Announce Type: cross Abstract: Large language models (LLMs) often exhibit performance disparities across languages, with naive multilingual fine-tuning frequently degrading performa

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

ConeSep: Cone-based Robust Noise-Unlearning Compositional Network for Composed Image Retrieval

DGX agent

arXiv:2604.20358v1 Announce Type: new Abstract: The Composed Image Retrieval (CIR) task provides a flexible retrieval paradigm via a reference image and modification text, but it heavily relies on exp

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Cooperative Profiles Predict Multi-Agent LLM Team Performance in AI for Science Workflows

DGX agent

arXiv:2604.20658v1 Announce Type: new Abstract: Multi-agent systems built from teams of large language models (LLMs) are increasingly deployed for collaborative scientific reasoning and problem-solvin

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

CRAFT: Training-Free Cascaded Retrieval for Tabular QA

DGX agent

arXiv:2505.14984v2 Announce Type: replace Abstract: Open-Domain Table Question Answering (TQA) involves retrieving relevant tables from a large corpus to answer natural language queries. Traditional d

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

CyberCertBench: Evaluating LLMs in Cybersecurity Certification Knowledge

DGX agent

arXiv:2604.20389v1 Announce Type: cross Abstract: The rapid evolution and use of Large Language Models (LLMs) in professional workflows require an evaluation of their domain-specific knowledge against

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Development and Preliminary Evaluation of a Domain-Specific Large Language Model for Tuberculosis Care in South Africa

DGX agent

arXiv:2604.19776v1 Announce Type: new Abstract: Tuberculosis (TB) is one of the world's deadliest infectious diseases, and in South Africa, it contributes a significant burden to the country's health

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

DialToM: A Theory of Mind Benchmark for Forecasting State-Driven Dialogue Trajectories

DGX agent

arXiv:2604.20443v1 Announce Type: cross Abstract: Large Language Models (LLMs) have been shown to possess Theory of Mind (ToM) abilities. However, it remains unclear whether this stems from robust rea

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Differentiable Conformal Training for LLM Reasoning Factuality

DGX agent

arXiv:2604.20098v1 Announce Type: new Abstract: Large Language Models (LLMs) frequently hallucinate, limiting their reliability in critical applications. Conformal Prediction (CP) addresses this by ca

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

DistortBench: Benchmarking Vision Language Models on Image Distortion Identification

DGX agent

arXiv:2604.19966v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in settings where sensitivity to low-level image degradations matters, including content moderatio

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Do Small Language Models Know When They're Wrong? Confidence-Based Cascade Scoring for Educational Assessment

DGX agent

arXiv:2604.19781v1 Announce Type: cross Abstract: Automated scoring of student work at scale requires balancing accuracy against cost and latency. In 'cascade' systems, small language models (LMs) han

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Dual Causal Inference: Integrating Backdoor Adjustment and Instrumental Variable Learning for Medical VQA

DGX agent

arXiv:2604.20306v1 Announce Type: cross Abstract: Medical Visual Question Answering (MedVQA) aims to generate clinically reliable answers conditioned on complex medical images and questions. However,

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Duluth at SemEval-2026 Task 6: DeBERTa with LLM-Augmented Data for Unmasking Political Question Evasions

DGX agent

arXiv:2604.20168v1 Announce Type: new Abstract: This paper presents the Duluth approach to SemEval-2026 Task 6 on CLARITY: Unmasking Political Question Evasions. We address Task 1 (clarity-level class

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Early-Stage Product Line Validation Using LLMs: A Study on Semi-Formal Blueprint Analysis

DGX agent

arXiv:2604.20523v1 Announce Type: cross Abstract: We study whether Large Language Models (LLMs) can perform feature model analysis operations (AOs) directly on semi-formal textual blueprints, i.e., co

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models

DGX agent

arXiv:2511.06209v4 Announce Type: replace Abstract: LLMs can solve complex tasks by generating long, multi-step reasoning chains. Test-time scaling (TTS) can further improve LLM performance by samplin

model-releasesarxiv-cs-ai
23 Apr 2026
← Previous
1…304305306307308…357
Next →