AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,603 results
27 May 2026

Cast a Wider Net: Coordinated Pass@K Policy Optimization for Code Reasoning

Model ReleasesDGX agent

arXiv:2605.27000v1 Announce Type: cross Abstract: Repeated sampling with a verifier is the standard way to allocate test-time compute for code generation, with pass@K as the canonical metric. Yet the

Causal Representation Learning for Generalisable Recommendation

Model ReleasesDGX agent

arXiv:2605.27043v1 Announce Type: cross Abstract: Predictive models trained on observational data often fail to generalise to the distributions they encounter when deployed, especially when the traini

Cesarean Scar Defect Segmentation in Transvaginal Ultrasound Images: a Dataset and Benchmark

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.26774v1 Announce Type: new Abstract: Cesarean Scar Defect (CSD) is one of the most prevalent complications following cesarean delivery. Transvaginal ultrasonography is widely used for prima

Chain Of Thought Compression: A Theoretical Analysis

Model ReleasesDGX agent

arXiv:2601.21576v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) has unlocked advanced reasoning abilities of Large Language Models (LLMs) with intermediate steps, yet incurs prohibitive com

ChartAct: A Benchmark for Dynamic Chart Understanding

Model ReleasesDGX agent

arXiv:2605.26994v1 Announce Type: new Abstract: Charts are widely used to present complex data for analysis and decision making. Existing chart understanding benchmarks mainly focus on static charts,

CIRCLED: A Multi-turn CIR Dataset with Consistent Dialogues across Domains

Model ReleasesDGX agent

arXiv:2605.26734v1 Announce Type: new Abstract: Existing Multi-Turn Composed Image Retrieval (MTCIR) datasets lack dialogue-history consistency and are restricted to the fashion domain. To address the

CktGen: Automated Analog Circuit Design with Generative Artificial Intelligence

Model ReleasesDGX agent

arXiv:2410.00995v3 Announce Type: replace Abstract: The automatic synthesis of analog circuits presents significant challenges. Most existing approaches formulate the problem as a single-objective opt

CleanSurvival: Automated data preprocessing for time-to-event models using reinforcement learning

Model ReleasesDGX agent

arXiv:2502.03946v5 Announce Type: replace Abstract: Data preprocessing is often paid little attention in machine learning, despite its potentially significant impact on model performance. While automa

Clinically-Grounded Counterfactual Reasoning for Medical Video Diagnosis

Model ReleasesDGX agent

arXiv:2605.26483v1 Announce Type: new Abstract: Medical video diagnosis involves inferring clinical decisions from dynamic tissue responses throughout examination processes. Existing methods rely on a

CNNs, Transformers, Hybrid, and Vision Language Models for Skin Cancer Detection

Model ReleasesDGX agent

arXiv:2605.26294v1 Announce Type: new Abstract: Skin cancer is a common and fast rising malignancy worldwide. Early detection is critical for improving outcomes. Deep learning models trained on dermos

CodecCap: High-Fidelity Codec-Inspired Residual Modeling for Dense Video Captioning

Model ReleasesDGX agent

arXiv:2605.26967v1 Announce Type: new Abstract: Existing video captioning methods struggle to balance visual fidelity and redundancy: holistic captions are compact but lose fine-grained evidence, wher

Cogent Security launches autonomous vulnerability response tools as AI-assisted exploits outpace scanners

Model ReleasesDGX agent

Cogent Security Inc., a startup that employs agentic artificial intelligence for vulnerability management, today launched two new platform capabilities aimed at compressing enterprise vulnerability re

Composition Collapse: Stable Factual Knowledge Does Not Imply Compositional Reasoning

Model ReleasesDGX agent

arXiv:2605.26789v1 Announce Type: new Abstract: Post-training is routinely evaluated through aggregate benchmark scores that treat multi-hop reasoning as a single capability -- as if a model that answ

Constraint acquisition needs better benchmarks

Model ReleasesDGX agent

arXiv:2605.26279v1 Announce Type: new Abstract: Constraint Acquisition (CA) and related research on the validation and enhancement of Mathematical Programming (MP) models from domain knowledge artifac

Constructing Industrial-Scale Optimization Modeling Benchmark

Model ReleasesDGX agent

arXiv:2602.10450v2 Announce Type: replace-cross Abstract: Optimization modeling underpins decision-making in logistics, manufacturing, energy, and finance, yet translating natural-language requirement

ConVer: Using Contracts and Loop Invariant Synthesis for Scalable Formal Software Verification

Model ReleasesDGX agent

arXiv:2605.27051v1 Announce Type: cross Abstract: Formal verification of large C programs is impeded by state-space explosion: Bounded Model Checking (BMC) tools must encode the entire state space up

COVD: Continual Open-Vocabulary Object Detection with Novel Concept Injection

Model ReleasesDGX agent

arXiv:2605.27116v1 Announce Type: new Abstract: Open-vocabulary object detection (OVD) has made significant progress, enabling detectors to generalize from seen to unseen categories. However, real-wor

Curation and Extraction of Drug-Related Entities from Reddit Platform

Model ReleasesDGX agent

arXiv:2605.26445v1 Announce Type: new Abstract: Physicians learn primarily about illicit drugs from clinical overdose cases, limiting their understanding of real-world usage. Meanwhile, drug users sha

Datacurve releases the DeepSWE coding benchmark, a 113-task test across 91 open-source repositories and five languages, and says GPT-5.5 is the leader at 70% (Michael Nuñez/VentureBeat)

Model ReleasesDGX agent

Michael Nuñez / VentureBeat: Datacurve releases the DeepSWE coding benchmark, a 113-task test across 91 open-source repositories and five languages, and says GPT-5.5 is the leader at 70% — For months,

Deep-layer limit and stability analysis of the basic forward-backward-splitting induced network (II): learning problems

Model ReleasesDGX agent

arXiv:2605.27133v1 Announce Type: cross Abstract: Deep unfolding neural networks derived from iterative optimization schemes and numerical ordinary/partial differential equations (ODEs/PDEs) have attr

DEI: Diversity in Evolutionary Inference for Quality-Diversity Search

Model ReleasesDGX agent

arXiv:2605.27130v1 Announce Type: cross Abstract: We present DEI: Diversity in Evolutionary Inference, a distributed Quality-Diversity (QD) search framework that assigns heterogeneous large language m

DelowlightSplat: Feed-Forward Gaussian Splatting for Lowlight 3D Scene Reconstruction

Model ReleasesDGX agent

arXiv:2605.26629v1 Announce Type: new Abstract: Novel-view synthesis and 3D reconstruction from sparse posed images are central to robotics and AR/VR. Yet, feed-forward 3D Gaussian reconstruction fail

Dense2MoE: Pushing the Pareto Frontier of On-Device LLMs via Unified Pruning and Upcycling

Model ReleasesDGX agent

arXiv:2605.26496v1 Announce Type: cross Abstract: The Mixture of Experts MoE architecture is highly promising for resource constrained on device deployments yet training these models from scratch incu

Developing a Totally Unimodular Linear Program for Optimal Conformance Checking: When and Why It Complements A*

Model ReleasesDGX agent

arXiv:2605.26938v1 Announce Type: new Abstract: Alignment-based conformance checking is the state-of-the-art approach for comparing observed process executions with normative process models. The stand

Device Context Protocol: A Compact, Safety-First Architecture for LLM-Driven Control of Constrained Devices

Model ReleasesDGX agent

arXiv:2605.26159v1 Announce Type: cross Abstract: Large language models are increasingly used as orchestrators of external tools via the Model Context Protocol (MCP), but MCP is built for software ser

DGLD: Domain-Gated Latent Diffusion for the Discovery of Novel Energetic Materials

Model ReleasesDGX agent

arXiv:2605.26540v1 Announce Type: cross Abstract: Energetic-materials performance gains translate directly into reduced propellant mass, smaller warheads, and more efficient civilian gas-generators, y

DIANOIA: Diagnostic Decomposition and Joint Optimization for Multi-Agent Reasoning

Model ReleasesDGX agent

arXiv:2602.08586v3 Announce Type: replace Abstract: Multi-agent LLM systems consistently outperform single-agent baselines, yet practitioners still cannot predict which design works for a new task or

Distribution-Aware Conformal Prediction: A Framework for generating efficient prediction intervals for time series

Model ReleasesDGX agent

arXiv:2605.26569v1 Announce Type: new Abstract: We present Distribution-aware Conformal Prediction (DCP), a unified framework integrating probabilistic predictors like Monte Carlo dropout, deep ensemb

Doppel launches agentic email security to disrupt phishing campaigns at the source

Model ReleasesDGX agent

Social engineering defense startup Doppel Inc. today launched Doppel Email Security, an agentic artificial intelligence layer that traces phishing messages back to attacker infrastructure and orchestr

Drive-P2D: A Progressive Perception-to-Decision Benchmark for VLMs in Autonomous Driving

Model ReleasesDGX agent

arXiv:2601.14702v2 Announce Type: replace Abstract: Autonomous driving requires reliable perception and safe decision-making in complex scenarios. Recent vision-language models (VLMs) demonstrate reas

DunbaaBERT: From Sacrifice to Semantics

Model ReleasesDGX agent

arXiv:2605.26935v1 Announce Type: new Abstract: Large language models have achieved strong performance across many NLP tasks, yet Urdu remains comparatively underexplored due to limited resources and

E3: Issue-Level Backtesting for Automated Research Critique

Model ReleasesDGX agent

arXiv:2605.27072v1 Announce Type: cross Abstract: We present E3, an automated review assistant that augments reviewers and engineering teams by identifying decision-relevant technical concerns in rese

ECHO-2: A Large-Scale Distributed Rollout Framework for Cost-Efficient Reinforcement Learning

Model ReleasesDGX agent

arXiv:2602.02192v5 Announce Type: replace Abstract: Reinforcement learning (RL) is a critical stage in post-training large language models (LLMs), involving repeated interaction between rollout genera

EconCausal: A Context-Aware Economic Reasoning Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2510.07231v4 Announce Type: replace-cross Abstract: Socio-economic causal effects depend heavily on their institutional and environmental contexts. The same intervention can produce different, e

ECSEL: Explainable Classification via Signomial Equation Learning

Model ReleasesDGX agent

arXiv:2601.21789v2 Announce Type: replace-cross Abstract: We introduce ECSEL, an explainable classification method that learns formal expressions in the form of signomial equations, motivated by the o

EdgeFlow: Edge-Map Augmented VLM-Based Flowchart Processing for Industrial Requirements Engineering

Model ReleasesDGX agent

arXiv:2605.27332v1 Announce Type: cross Abstract: Flowcharts are widely used in industrial requirements, but usually remain embedded as static images. Vision Language Models (VLMs) show promise in the

Efficient Prediction of SO(3)-Equivariant Hamiltonian Matrices via SO(2) Local Frames

Model ReleasesDGX agent

arXiv:2506.09398v3 Announce Type: replace Abstract: We consider the task of predicting Hamiltonian matrices to accelerate electronic structure calculations, which plays an important role in physics, c

EgoProx: Evaluating MLLMs on Egocentric 3D Proximity Reasoning Across a Cognitive Hierarchy

Model ReleasesDGX agent

arXiv:2605.24456v2 Announce Type: replace Abstract: Humans constantly reason about 3D proximity, the relations between their body and surrounding objects, to guide perception and action in daily life.

EHRSummarizer: A Privacy-Aware, FHIR-Native Reference Architecture for Source-Grounded EHR Summarization

Model ReleasesDGX agent

arXiv:2601.01668v2 Announce Type: replace-cross Abstract: Clinicians routinely navigate fragmented electronic health record (EHR) interfaces to assemble a coherent picture of a patient's problems, med

Election information and safeguards in 2026

Model ReleasesDGX agent

OpenAI outlines measures and policies designed to protect the integrity of the 2026 election cycle, likely addressing how their AI systems will handle election-related information and misinformation.

Enhancing Autonomous Online Intrusion Detection for IoT with Balanced Learning, Reliable Pseudo-Labels, and Lightweight Architectures

Model ReleasesDGX agent

arXiv:2605.26166v1 Announce Type: cross Abstract: The rapid proliferation of Internet of Things (IoT) devices has created an urgent demand for adaptive, resource-efficient Intrusion Detection Systems

ENPMR-Bench: Benchmarking Proactive Memory Retrieval for Emotional Support Agents

Model ReleasesDGX agent

arXiv:2605.27240v1 Announce Type: new Abstract: Memory-augmented language agents are increasingly deployed in affective applications such as emotional support, where understanding and responding to us

Entropy Sentinel: Continuous LLM Accuracy Monitoring from Decoding Entropy Traces in STEM

Model ReleasesDGX agent

arXiv:2601.09001v4 Announce Type: replace Abstract: Deploying LLMs raises two coupled challenges: (1) monitoring--estimating where a model underperforms as traffic and domains drift--and (2) improveme

EpiCurveBench: Evaluating VLMs on Epidemic Curve Digitization

Model ReleasesDGX agent

arXiv:2605.27195v1 Announce Type: new Abstract: Chart-to-data extraction with vision-language models (VLMs) is increasingly evaluated on benchmarks that show diminishing headroom (frontier VLMs exceed

EpiQAL: Benchmarking Large Language Models in Epidemiological Question Answering and Reasoning

Model ReleasesDGX agent

arXiv:2601.03471v3 Announce Type: replace-cross Abstract: Reliable epidemiological reasoning requires synthesizing study evidence to infer disease burden, transmission dynamics, and intervention effec

Evi-Steer: Learning to Steer Biomedical Vision-Language Models through Efficient and Generalizable Evidential Tuning

Model ReleasesDGX agent

arXiv:2605.26292v1 Announce Type: cross Abstract: Parameter-efficient adaptation of vision-language foundation models is crucial for precise multimodal understanding of biomedical images, yet existing

Exclusive: Unravel Data launches autonomous optimization engine for Databricks, Snowflake and BigQuery

Model ReleasesDGX agent

Unravel Data Systems Inc. is expanding beyond observability and FinOps software with a new autonomous optimization engine designed to automatically tune and remediate enterprise data platforms running

Exploiting Local Dynamics Regularity for Reusable Skills in Offline Hierarchical RL

Model ReleasesDGX agent

arXiv:2605.26371v1 Announce Type: new Abstract: Hierarchical Reinforcement Learning (HRL) promises to solve long-horizon Reinforcement Learning (RL) tasks more efficiently than non-hierarchical counte

Extra-Merge: Tracing the Rank-1 Subspace of Model Merging in Language Model Pre-Training

Model ReleasesDGX agent

arXiv:2605.26484v1 Announce Type: new Abstract: Model merging has emerged as a lightweight paradigm for enhancing Large Language Models (LLMs), yet its underlying mechanisms remain poorly understood.

FAB-Bench: A Framework for Adaptive RAG Benchmarking in Semiconductor Manufacturing

Model ReleasesDGX agent

arXiv:2605.26476v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become critical for knowledge-intensive applications, yet evaluating its performance in vertical domains remain

Faithfulness Evaluation for Decoder-only LLM Attributions with Controlled Retained Information

Model ReleasesDGX agent

arXiv:2601.03089v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly evaluated with input attribution methods, yet comparing such explanations remains challenging. E

Falcon-X: A Time Series Foundation Model for Heterogeneous Multivariate Modeling

Model ReleasesDGX agent

arXiv:2605.27286v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) are transforming the forecasting paradigm through large-scale cross-domain pretraining. However, most existing T

Fast, faster, Qwen. 🚀 Thrilled to see Qwen3.5 reaching a record-breaking 580 tps for agentic workloads on the TokenSpeed engine! This miles…

Model ReleasesDGX agent

Fast, faster, Qwen. 🚀 Thrilled to see Qwen3.5 reaching a record-breaking 580 tps for agentic workloads on the TokenSpeed engine! This milestone wouldn't be possible without our incredible partners. Hu

FedTreeLoRA: Reconciling Statistical and Functional Heterogeneity in Federated LoRA Fine-Tuning

Model ReleasesDGX agent

arXiv:2603.13282v2 Announce Type: replace-cross Abstract: Federated Learning (FL) with Low-Rank Adaptation (LoRA) has become a standard for privacy-preserving LLM fine-tuning. However, existing person

FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies

Model ReleasesDGX agent

arXiv:2605.27284v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are increasingly expected to not only complete robot tasks, but also follow human instructions about how those tas

Focal Reward: Balanced Reinforcement Learning under Rubric-Based Rewards

Model ReleasesDGX agent

arXiv:2605.26579v1 Announce Type: new Abstract: The open-ended generation in LLMs usually requires multi-dimensional rubrics to adequately assess quality and guide the improvement of reinforcement lea

From PDF to RAG-Ready: Evaluating Document Conversion Frameworks for Domain-Specific Question Answering

Model ReleasesDGX agent

arXiv:2604.04948v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems depend critically on the quality of document preprocessing, yet no prior study has evaluated PDF

GEM: Geometric Entropy Mixing for Optimal LLM Data Curation

Model ReleasesDGX agent

arXiv:2605.26121v1 Announce Type: cross Abstract: LLM pre-training efficacy increasingly depends on data composition rather than sheer volume. Yet, optimal mixing is hindered by categorization flaws:

Gemini Embedding 2: A Native Multimodal Embedding Model from Gemini

Model ReleasesDGX agent

arXiv:2605.27295v1 Announce Type: new Abstract: We introduce Gemini Embedding 2, a native multimodal embedding model that allows embedding video, audio, image, and text modalities in a unified represe

GeoFaith: A Spatio-Temporal Dual View of Faithful Chain-of-Thought

Model ReleasesDGX agent

arXiv:2605.26893v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) reasoning has advanced large language models (LLMs), but outcome-based supervision leads to pervasive post-hoc rationalization,

← Previous
1…204205206207208…377
Next →