AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Human-LLM Dialogue Improves Diagnostic Accuracy in Emergency Care

DGX agent

arXiv:2605.08533v1 Announce Type: new Abstract: Clinical decision-making in emergency medicine demands rapid, accurate diagnoses under uncertainty. Despite benchmark progress, evidence for LLMs as int

model-releasesarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Hunyuan3D 2.0: Scaling Diffusion Models for High Resolution Textured 3D Assets Generation

DGX agent

arXiv:2501.12202v4 Announce Type: replace Abstract: We present Hunyuan3D 2.0, an advanced large-scale 3D synthesis system for generating high-resolution textured 3D assets. This system includes two fo

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

HyperSpace: A Generalized Framework for Spatial Encoding in Hyperdimensional Representations

DGX agent

arXiv:2604.15113v2 Announce Type: replace Abstract: Vector Symbolic Architectures (VSAs) provide a well-defined algebraic framework for compositional representations in hyperdimensional spaces. We int

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Hystar: Hypernetwork-driven Style-adaptive Retrieval via Dynamic SVD Modulation

DGX agent

arXiv:2605.10009v1 Announce Type: new Abstract: Query-based image retrieval (QBIR) requires retrieving relevant images given diverse and often stylistically heterogeneous queries, such as sketches, ar

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Identifying Backdoored Graphs in Graph Neural Network Training: An Explanation-Based Approach with Novel Metrics

DGX agent

arXiv:2403.18136v3 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) have gained popularity in numerous domains, yet they are vulnerable to backdoor attacks that can compromise their

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Identifying Multi-Hit Cancer Drivers Without Massive Parallelization: A CP, MIP, and Column Generation Framework

DGX agent

arXiv:2602.22551v2 Announce Type: replace-cross Abstract: Cancer is often driven by specific combinations of an estimated two to nine gene mutations, known as multi-hit combinations. Identifying these

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Illusion-Aware Visual Preprocessing and Anti-Illusion Prompting for Classic Illusion Understanding in Vision-Language Models

DGX agent

arXiv:2605.08841v1 Announce Type: new Abstract: Vision-Language Models (VLMs) exhibit systematic bias toward visual illusions, recalling memorized facts rather than perceiving actual visual difference

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Improving Generalization by Permutation Routing Across Model Copies

DGX agent

arXiv:2605.09256v1 Announce Type: cross Abstract: We introduce a use of the (M)-cover (or (M)-layer) transform for machine learning. The method replicates a model (M) times, but instead of coupling th

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Improving TMS EEG Signal Quality for Closed-Loop Neuro Stimulation via Source-Domain Denoising

DGX agent

arXiv:2605.08184v1 Announce Type: cross Abstract: This research addresses a validated TMS EEG cleaning pipeline and a corresponding benchmark dataset. It evaluates two widely used artifact removal pip

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

In-Context Fixation: When Demonstrated Labels Override Semantics in Few-Shot Classification

DGX agent

arXiv:2605.08295v1 Announce Type: cross Abstract: While random demonstration labels barely hurt in-context learning (Min et al., 2022), we show that homogeneous labels--even semantically valid ones--c

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Incremental Multilingual Text2Cypher with Adapter Combination

DGX agent

arXiv:2601.16097v2 Announce Type: replace Abstract: Large Language Models enable users to access database using natural language interfaces using tools like Text2SQL, Text2SPARQL, and Text2Cypher, whi

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

IndustryBench: Probing the Industrial Knowledge Boundaries of LLMs

DGX agent

arXiv:2605.10267v1 Announce Type: new Abstract: In industrial procurement, an LLM answer is useful only if it survives a standards check: recommended material must match operating condition, every par

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Instruction Adherence in Coding Agent Configuration Files: A Factorial Study of Four File-Structure Variables

DGX agent

arXiv:2605.10039v1 Announce Type: cross Abstract: Frontier coding agents read configuration files (CLAUDE.md, AGENTS.md, Cursor Rules) at session start and are expected to follow the conventions insid

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Intervention-Based Time Series Causal Discovery via Simulator-Generated Interventional Distributions

DGX agent

arXiv:2605.09870v1 Announce Type: cross Abstract: We propose SVAR-FM (Structural VAR with Flow Matching), a framework for time series causal discovery that treats a physics-based simulator as a mechan

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Inverse Design of Multi-Layer Sub-Pixel-Resolution RF Passives Through Grayscale Diffusion with Flexible S-Parameter Conditioning

DGX agent

arXiv:2605.08233v1 Announce Type: cross Abstract: Inverse design of RF passive components from S-parameters is a high-dimensional, ill-posed problem, and prior generative approaches are limited to sin

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

IPAD-CLIP: Teaching CLIP to Detect Image Local Perceptual Artifacts

DGX agent

arXiv:2605.08664v1 Announce Type: new Abstract: Current image quality assessment methods are heavily biased towards global distortions (e.g., noise, blur), neglecting local perceptual artifacts such a

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Is Your Driving World Model an All-Around Player?

DGX agent

arXiv:2605.10858v1 Announce Type: new Abstract: Today's driving world models can generate remarkably realistic dash-cam videos, yet no single model excels universally. Some generate photorealistic tex

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

ITLC at SemEval-2026 Task 11: Normalization and Deterministic Parsing for Formal Reasoning in LLMs

DGX agent

arXiv:2603.02676v2 Announce Type: replace-cross Abstract: Large language models suffer from content effects in reasoning tasks, particularly in multi-lingual contexts. We introduce a novel method that

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

jina-embeddings-v5-omni: Text-Geometry-Preserving Multimodal Embeddings via Frozen-Tower Composition

DGX agent

arXiv:2605.08384v1 Announce Type: new Abstract: In this work, we introduce frozen-encoder model composition, a novel approach to multimodal embedding models. We build on the VLM-style architecture, in

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

jNO: A JAX Library for Neural Operator and Foundation Model Training

DGX agent

arXiv:2605.10159v1 Announce Type: new Abstract: jNO (jax Neural Operators) is a JAX-native library for neural operators and foundation models with unified support for both data-driven and physics-info

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

JODA: Composable Joint Dynamics for Articulated Objects

DGX agent

arXiv:2605.09954v1 Announce Type: cross Abstract: Articulated objects used in simulation and embodied AI are typically specified by geometry and kinematic structure, but lack the fine-grained dynamica

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

K12-KGraph: A Curriculum-Aligned Knowledge Graph for Benchmarking and Training Educational LLMs

DGX agent

arXiv:2605.09635v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in K-12 education, yet existing benchmarks such as C-Eval, CMMLU, GaokaoBench, and EduEval mainly eva

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

KAN Text to Vision? The Exploration of Kolmogorov-Arnold Networks for Multi-Scale Sequence-Based Pose Animation from Sign Language Notation

DGX agent

arXiv:2605.09572v1 Announce Type: cross Abstract: Sign language production from symbolic notation offers a scalable route to accessible sign animation. We present KANMultiSign, a multi-scale sequence

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

KARMA-MV: A Benchmark for Causal Question Answering on Music Videos

DGX agent

arXiv:2605.08175v1 Announce Type: cross Abstract: While significant progress has been made in Video Question Answering and cross-modal understanding, causal reasoning about how visual dynamics drive m

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

KEPIL: Knowledge-Enhanced Prompt-Image Learning for Prompt-Robust Disease Detection

DGX agent

arXiv:2605.09132v1 Announce Type: new Abstract: Vision--language models (VLMs) show promise for clinical decision support in radiology because they enable joint reasoning over radiological images and

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Knowledge is Not Enough: Injecting RL Skills for Continual Adaptation

DGX agent

arXiv:2601.11258v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) face the 'knowledge cutoff' challenge, where their frozen parametric memory prevents direct internalization of ne

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

KV-RM: Regularizing KV-Cache Movement for Static-Graph LLM Serving

DGX agent

arXiv:2605.09735v1 Announce Type: cross Abstract: Static-graph LLM decoders provide predictable launches, fixed tensor shapes, and low submission overhead, but online decoding exposes highly irregular

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Language Models Without a Trainable Input Embedding Table: Learning from Fixed Minimal Binary Token Codes

DGX agent

arXiv:2605.09751v1 Announce Type: new Abstract: Trainable input embedding tables are a standard component of modern language models. We ask whether they are actually necessary at the input interface.

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Large Language Models as Students Who Think Aloud: Overly Coherent, Verbose, and Confident

DGX agent

arXiv:2602.01015v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly embedded in AI-based tutoring systems. Can they faithfully model novice reasoning and metacognitive ju

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Latent Geometry Beyond Search: Amortizing Planning in World Models

DGX agent

arXiv:2605.08732v1 Announce Type: cross Abstract: Modern vision-based world models can represent observations as compact yet expressive latent manifolds, but fast goal-oriented planning in these space

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Lattice Deduction Transformers

DGX agent

arXiv:2605.08605v1 Announce Type: cross Abstract: We introduce the Lattice Deduction Transformer (LDT), a recurrent transformer that approximates logically sound deduction by projecting its latent sta

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Layer Collapse in Diffusion Language Models

DGX agent

arXiv:2605.06366v2 Announce Type: replace Abstract: Diffusion language models (DLMs) have recently emerged as competitive alternatives to autoregressive (AR) language models, yet differences in their

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

LEAD: Length-Efficient Adaptive and Dynamic Reasoning for Large Language Models

DGX agent

arXiv:2605.09806v1 Announce Type: cross Abstract: Large reasoning models, such as OpenAI o1 and DeepSeek-R1, tend to become increasingly verbose as their reasoning capabilities improve. These inflated

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

LEAF-SQL: Level-wise Exploration with Adaptive Fine-graining for Text-to-SQL Skeleton Prediction

DGX agent

arXiv:2605.09295v1 Announce Type: new Abstract: Text-to-SQL translates natural language questions into executable SQL queries, enabling intuitive database access for non-experts. While large language

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Learning Agile Striker Skills for Humanoid Soccer Robots from Noisy Sensory Input

DGX agent

arXiv:2512.06571v3 Announce Type: replace Abstract: Learning fast and robust ball-kicking skills is a critical capability for humanoid soccer robots, yet it remains a challenging problem due to the ne

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

Learning Confidence Ellipsoids and Applications to Robust Subspace Recovery

DGX agent

arXiv:2512.16875v4 Announce Type: replace-cross Abstract: We study the problem of finding confidence ellipsoids for an arbitrary distribution in high dimensions. Given samples from a distribution D an

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Learning from Trials and Errors: Reflective Test-Time Planning for Embodied LLMs

DGX agent

arXiv:2602.21198v2 Announce Type: replace-cross Abstract: Embodied LLMs endow robots with high-level task reasoning, but they cannot reflect on what went wrong or why, turning deployment into a sequen

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Learning Less Is More: Premature Upper-Layer Attention Specialization Hurts Language Model Pretraining

DGX agent

arXiv:2605.10504v1 Announce Type: new Abstract: A causal-decoder block is hierarchical: lower layers build the residual basis that upper layers attend over. We identify a failure mode in GPT pretraini

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Learning Multi-Indicator Weights for Data Selection: A Joint Task-Model Adaptation Framework with Efficient Proxies

DGX agent

arXiv:2605.09665v1 Announce Type: cross Abstract: Data selection is a key component of efficient instruction tuning for large language models, as recent work has shown that data quality often matters

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Learning to Perceive 'Where': Spatial Pretext Tasks for Robust Self-Supervised Learning

DGX agent

arXiv:2605.09963v1 Announce Type: new Abstract: Existing self-supervised learning (SSL) methods primarily learn object-invariant representations but often neglect the spatial structure and relationshi

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Learning to Sparsify Stochastic Linear Bandits

DGX agent

arXiv:2605.10151v1 Announce Type: new Abstract: This paper addresses the problem of learning to sparsify stochastic linear bandits, where a decision-maker sequentially selects actions from a high-dime

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

LegalCiteBench: Evaluating Citation Reliability in Legal Language Models

DGX agent

arXiv:2605.10186v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into legal drafting and research workflows, where incorrect citations or fabricated precedent

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Less Diverse, Less Safe: The Indirect But Pervasive Risk of Test-Time Scaling in Large Language Models

DGX agent

arXiv:2510.08592v3 Announce Type: replace-cross Abstract: Test-Time Scaling (TTS) improves LLM reasoning by exploring multiple candidate responses and then operating over this set to find the best out

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

LEVI: Stronger Search Architectures Can Substitute for Larger LLMs in Evolutionary Search

DGX agent

arXiv:2605.09764v1 Announce Type: cross Abstract: LLM-guided evolutionary methods such as AlphaEvolve have proven effective in domains like math, systems research, and algorithmic discovery, but their

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

LightAVSeg: Lightweight Audio-Visual Segmentation

DGX agent

arXiv:2605.08805v1 Announce Type: new Abstract: Audio-Visual Segmentation (AVS) targets pixel level localization of sounding emitting objects in videos. However, existing models rely on dense cross-mo

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities

DGX agent

arXiv:2605.10810v1 Announce Type: new Abstract: We introduce an automatically generated benchmark for predicting hidden text in technical papers. A paper supplies visible context X and a hidden contin

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

LimeCross: Context-Conditioned Layered Image Editing with Structural Consistency

DGX agent

arXiv:2605.10319v1 Announce Type: new Abstract: Layered image assets are widely used in real-world creative workflows, enabling non-destructive iteration and flexible re-composition. Recent advances i

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

LiteMedCoT-VL: Parameter-Efficient Adaptation for Medical Visual Question Answering

DGX agent

arXiv:2605.09384v1 Announce Type: cross Abstract: The reasoning gap between large and compact vision-language models (VLMs) limits the deployment of medical AI on portable clinical devices. Compact VL

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…255256257258259…361
Next →