AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

CoCoSI: Collaborative Cognitive Map Construction for Spatial Intelligence

DGX agent

arXiv:2606.10401v1 Announce Type: new Abstract: Spatial intelligence is a key frontier for multimodal large language models (MLLMs), enabling them to reason about the physical world from visual experi

model-releasesarxiv-cs-cv
10 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CodeAlchemy: Synthetic Code Rewriting at Scale

DGX agent

arXiv:2606.10087v1 Announce Type: new Abstract: Pre-training on raw code teaches syntax but provides sparse signal for diverse real-world task formats. While synthetic data has proven transformative f

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

CollabSkill: Evaluating Human-Agent Collaboration On Real-World Tasks

DGX agent

arXiv:2606.09833v1 Announce Type: cross Abstract: AI agents are reshaping the workspace, leading to drastic change of how humans work. Despite the considerable potential of human-agent collaboration b

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

ComBench: A Benchmark for Rigorous Proof Reasoning and Constructive Realization in Olympiad-Level Combinatorics

DGX agent

arXiv:2606.10479v1 Announce Type: new Abstract: Combinatorics is central to Olympiad-level mathematical problem solving, requiring deep discrete reasoning, creative constructions, and rigorous structu

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data

DGX agent

arXiv:2601.18026v2 Announce Type: replace Abstract: Language identification (LID) is a fundamental step in curating multilingual corpora. However, LID models still perform poorly for many languages, e

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Compile Once, Differentiate Everywhere: A Differentiable Meta-Circular Interpreter

DGX agent

arXiv:2606.09930v1 Announce Type: cross Abstract: The boundary between program execution and gradient-based optimization has long limited the use of code itself as a learnable scientific model. We pre

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Constructing coherent spatial memory in LLM agents through graph rectification

DGX agent

arXiv:2510.04195v2 Announce Type: replace Abstract: Given a map description through global traversal navigation instructions, an LLM can often infer the implicit spatial layout and answer user queries

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Continual LLM Upcycling: A Predictor-Gated Bank-Wise Sparsity Training Recipe for Dense-to-Sparse LLMs

DGX agent

arXiv:2606.10722v1 Announce Type: new Abstract: We study dense-to-sparse continual training as a way to construct channel-sparse large language models from dense checkpoints. Starting from a Qwen2.5-8

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Continuous Neural Reparameterization as a Deep Geometric Prior for Robust Fixed-Chart UV Repair

DGX agent

arXiv:2606.10050v1 Announce Type: cross Abstract: Traditional UV unwrapping relies on direct optimization of geometric distortion energies and can fail through invalid initialization, local minima, or

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

ConvMemory v2: A Recall-Preserving Top-10 Evidence Reranker for Conversational Memory Retrieval

DGX agent

arXiv:2606.10842v1 Announce Type: new Abstract: We describe ConvMemory v2, an opt-in token-evidence reranker that sits after the lightweight ConvMemory v1 reranker and reorders only v1's protected top

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

CoTAL: Human-in-the-Loop Prompt Engineering for Generalizable Formative Assessment Scoring and Feedback

DGX agent

arXiv:2504.02323v4 Announce Type: replace Abstract: Large language models (LLMs) have created new opportunities to assist teachers and support student learning. While researchers have explored various

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Cyst-X: A Multi-Center MRI Benchmark and Federated Learning Framework for Malignancy-Risk Stratification of Pancreatic Cystic Neoplasm

DGX agent

arXiv:2507.22017v4 Announce Type: replace-cross Abstract: Pancreatic cancer is projected to be the second-deadliest cancer by 2030, making early detection critical. Intraductal papillary mucinous neop

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Data-Driven Dynamic Assortment in Online Platforms: Learning about Two Sides

DGX agent

arXiv:2606.11118v1 Announce Type: new Abstract: We study a dynamic assortment problem on a two-sided service platform with incomplete information and heterogeneous customers in a discrete-time setting

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Data-Driven Runway and Taxiway Exits Prediction of Landing Aircraft: A Case Study at Hartsfield-Jackson Atlanta International Airport

DGX agent

arXiv:2606.11017v1 Announce Type: new Abstract: Airport surface operations increasingly constrain performance at high-throughput hubs. This study examines arrival taxi-in decisions at Hartsfield-Jacks

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

DB-3DME: From Dataset to Benchmark for Human-aligned Automatic 3D Mesh Evaluation

DGX agent

arXiv:2606.10142v1 Announce Type: new Abstract: Recent advances in 3D generation have led to substantial improvements in realism, controllability, and efficiency, yet the evaluation of 3D assets remai

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Deployment-Time Memorization in Foundation-Model Agents

DGX agent

arXiv:2606.10062v1 Announce Type: new Abstract: Foundation-model agents are increasingly long-lived systems that remember users across interactions, making memorization an explicit deployment-time fun

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Detecting Knowledge Gaps from Conversational AI Interactions Using Curriculum Prerequisite Graphs

DGX agent

arXiv:2606.10736v1 Announce Type: cross Abstract: Large online courses generate thousands of student questions directed at conversational AI teaching assistants, yet these interaction logs remain larg

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Discovering Interpretable Multi-Parameter Control Policies for Evolutionary Algorithms Using Deep Reinforcement Learning

DGX agent

arXiv:2606.10129v1 Announce Type: new Abstract: While deep Reinforcement Learning (deep-RL) has been increasingly applied to parameter control in evolutionary algorithms, rigorous theoretical analysis

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Disjoint or Overlapping? Inference Windowing for Reconstruction-Based Time Series Anomaly Detection

DGX agent

arXiv:2606.09874v1 Announce Type: new Abstract: Reconstruction-based methods are widely used for time series anomaly detection, where models are trained to reconstruct subsequences, and anomalies are

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Divide-and-Conquer Modeling for the CTF-4-Science Lorenz Benchmark

DGX agent

arXiv:2606.10084v1 Announce Type: cross Abstract: This work presents a divide-and-conquer modeling strategy for the CTF-4-Science Lorenz benchmark, which evaluates chaotic-system prediction across twe

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Divide and Cooperate: Role-Decomposed Multi-Agent LLM Training with Cross-Agent Learning Signals

DGX agent

arXiv:2606.10684v1 Announce Type: cross Abstract: Modern language agents which perform multi-step reasoning have shown strong performance in knowledge-intensive question answering. However, existing a

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Do Vision-Language Models See or Guess? Measuring and Reducing Textual-Prior Reliance with a Phrasing-Controlled Benchmark

DGX agent

arXiv:2606.10400v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed where answers must follow from what is in the image, yet they often answer from textual priors,

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Do VLMs Reason Like Engineers? A Benchmark and a Stage-wise Evaluation

DGX agent

arXiv:2606.10833v1 Announce Type: new Abstract: Vision-Language Models (VLMs) demonstrate strong performance on general multimodal reasoning benchmarks, yet their ability to perform engineering reason

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Domain Adapted Large Language Models for Additive Manufacturing

DGX agent

arXiv:2603.22017v2 Announce Type: replace Abstract: This work presents a collection of multi-modal domain adapted large language models built upon the instruction tuned variants of open weight models

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Don't waste SAM

DGX agent

arXiv:2606.10696v1 Announce Type: new Abstract: Meta AI has recently released the Segment Anything Model (SAM), which demonstrates exceptional zero-shot image segmentation performance across various t

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Dual-Branch Gated Fusion for Open-Set Audio Deepfake Source Tracing

DGX agent

arXiv:2606.10223v1 Announce Type: cross Abstract: Attributing a synthetic utterance to its originating system remains an open challenge: closed-set models fail to reject unseen synthesizers and produc

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Earth-OneVision: Extending Remote Sensing Multimodal Large Language Models to More Sensor Modalities and Tasks

DGX agent

arXiv:2606.10819v1 Announce Type: cross Abstract: RS-MLLMs enable natural-language understanding and spatial reasoning over earth observation imagery. However, existing models support only a narrow ra

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents

DGX agent

arXiv:2606.11182v1 Announce Type: cross Abstract: In this paper, we propose EEVEE, the first multi-dataset test-time prompt learning framework for LLM agents, enabling test-time prompt learning under

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Effective Reinforcement Learning for Agentic Search by Recycling Zero-Variance Queries During Training

DGX agent

arXiv:2606.10709v1 Announce Type: cross Abstract: The use of GRPO-style algorithms has become the standard strategy for training LLM search agents under outcome-only rewards. With these algorithms, a

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Effective Training Principles of Physical Reservoirs

DGX agent

arXiv:2606.10130v1 Announce Type: cross Abstract: Reservoir computers benefit from the inherent complexity of optical phenomena, which provide rich, often nonlinear dynamics. However, training directl

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Efficient RWKV-based Representation Learning for 3D Point Clouds

DGX agent

arXiv:2606.10395v1 Announce Type: new Abstract: The recent receptance weighted key value (RWKV) model combines RNN-style recurrence, offering a linear-complexity alternative to Transformers' quadratic

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Efficient-WAM: A 1B-Parameter World-Action Model with Low-Cost Future Imagination

DGX agent

arXiv:2606.10040v1 Announce Type: new Abstract: World-Action Models (WAMs) have emerged as a promising paradigm for embodied control by coupling future visual prediction with action generation. Howeve

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

Enabling Progressive Whole-slide Image Analysis with Multi-scale Pyramidal Network

DGX agent

arXiv:2602.01951v2 Announce Type: replace Abstract: Multiple-instance Learning (MIL) is commonly used for computational pathology (CPath), where multi-scale features are essential for capturing both f

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Evaluating Research-Level Math Proofs via Strict Step-Level Verification

DGX agent

arXiv:2606.10799v1 Announce Type: new Abstract: Large Language Models (LLMs) struggle to rigorously verify complex mathematical proofs. Standard global evaluation approaches suffer from 'context poiso

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Expert-Level Crisis Detection in Mental Health Conversations

DGX agent

arXiv:2606.10380v1 Announce Type: cross Abstract: Real-world crisis intervention is inherently conversational, yet existing research largely focuses on static texts.Real-world crisis intervention is i

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Exploratory Responsiveness and Adaptive Rigidity under AI-Assisted Optimization

DGX agent

arXiv:2606.10086v1 Announce Type: new Abstract: This paper develops a theory of exploratory adaptation under AI-assisted optimization. The central argument is that the long-run adaptive effects of AI

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Fact-Augmented Lookahead Planning for LLM Agents

DGX agent

arXiv:2506.09171v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly capable, but LLM agents still struggle to plan effectively in interactive, partially observable,

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

FADA: Accessible fetal ultrasound interpretation and annotation with a selectively distilled unified vision-language model

DGX agent

arXiv:2606.11106v1 Announce Type: cross Abstract: A global shortage of trained sonographers limits prenatal ultrasound screening in low- and middle-income countries, where over half of pregnant women

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Failure Modes of Deep Multi-Agent RL in Asynchronous Pricing: Reproducible Triggers, Trace Diagnostics, and a Partial Fix

DGX agent

arXiv:2606.09884v1 Announce Type: cross Abstract: We study two reproducible failure modes of deep multi-agent reinforcement learning in continuous-time pricing markets: (i) tacit cartel formation betw

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

FailureScope: Cross-Regime Behavioral Diagnosis of Language Model Weaknesses

DGX agent

arXiv:2606.09878v1 Announce Type: new Abstract: Standard benchmarks report aggregate accuracy, but practitioners need to know which specific capabilities a model lacks. We introduce FailureScope, a be

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Fisher-Guided Progressive Parameter Selection for Adaptive Fine-Tuning

DGX agent

arXiv:2606.10196v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) aims to adapt pretrained models with a small trainable parameter subset, however, most existing methods choose

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Flaws in the LLM Automation Narrative

DGX agent

arXiv:2606.11166v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly described as performing at the level of human experts on knowledge economy tasks. These claims are prima

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

FreshRetailNet-LT: A Stockout-Annotated Censored Demand Dataset for Latent Demand Recovery and Forecasting in Fresh Retail

DGX agent

arXiv:2505.16319v3 Announce Type: replace Abstract: Accurate demand estimation is critical for the retail business in guiding the inventory and pricing policies of perishable products. However, it fac

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

From Observation to Intervention: A Causal Audit of Expert Importance in Mixture-of-Experts Models

DGX agent

arXiv:2606.10703v1 Announce Type: cross Abstract: Interpretability methods routinely use population-level summary statistics over observed model behaviour to license claims about the effects of target

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

From Patches to Patients: A study of the tile-to-slide performance transferability in Digital Pathology

DGX agent

arXiv:2606.10778v1 Announce Type: new Abstract: Foundation Models (FMs) have recently redefined the state-of-the-art in histopathology by providing robust representations for whole-slide image (WSI) a

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Frontier Coding Agents Use Metaprogramming to Adapt to Unfamiliar Programming Languages

DGX agent

arXiv:2606.10933v1 Announce Type: new Abstract: LLM-based coding agents are usually evaluated in familiar software settings: mainstream languages, common libraries, and public repositories. These benc

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Gaming AI-Assisted Peer Reviews Poses New Risks to the Scientific Community

DGX agent

arXiv:2606.10159v1 Announce Type: cross Abstract: AI is increasingly used to support scientific peer review, from manuscript screening, reviewer assistance to editorial triage. Although such systems p

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Geometry-Aware Reinforcement Learning for 2D Irregular Nesting

DGX agent

arXiv:2606.10611v1 Announce Type: cross Abstract: Traditional heuristic solvers for the 2D irregular nesting problem share a fundamental limitation: they are blind to polygon geometry, relying on guid

model-releasesarxiv-cs-cv
10 Jun 2026
← Previous
1…136137138139140…361
Next →