AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,628 results
Model Releases

Distributionally Robust Transfer Learning with Structurally Missing Covariates, with Application to Cross-National Cardiac Arrest Prediction

DGX agent

arXiv:2605.24212v1 Announce Type: cross Abstract: Deploying clinical prediction models across healthcare systems often fails when key training covariates are unavailable at deployment and labeled outc

model-releasesarxiv-cs-ai
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

DIVER-1: Scaling Intracranial EEG Foundation Models for Transferable Representations

DGX agent

arXiv:2512.19097v3 Announce Type: replace-cross Abstract: Intracranial EEG (iEEG) provides direct, millisecond-scale recordings of human neural activity, but reusable representation learning is diffic

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Double Triangle Annotation: A Scalable Human-in-the-Loop Framework for High-Precision Historical Document Annotation

DGX agent

arXiv:2605.25781v1 Announce Type: new Abstract: Evaluating structured-information extraction from historical documents at scale requires high-precision ground-truth annotations, yet traditional manual

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

DRInQ: Evaluating Conversational Implicature with Controlled Context Variation

DGX agent

arXiv:2605.24267v1 Announce Type: new Abstract: Human conversation relies heavily on conversational implicature, in which speakers convey meanings that are suggested rather than explicitly stated. Alt

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

DropoutTS: Sample-Adaptive Dropout for Robust Time Series Forecasting

DGX agent

arXiv:2601.21726v2 Announce Type: replace Abstract: Deep time series models are vulnerable to noisy data ubiquitous in real-world applications. Existing robustness strategies either prune data or rely

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

DRScaffold: Boosting Dense-Scene Reasoning in Lightweight Vision Language Models

DGX agent

arXiv:2605.26038v1 Announce Type: cross Abstract: Lightweight vision-language models perform competitively on standard benchmarks yet fail systematically in dense-scene reasoning, where multiple objec

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Dynamics Reveals Structure: Challenging the Linear Propagation Assumption

DGX agent

arXiv:2601.21601v2 Announce Type: replace-cross Abstract: Neural networks adapt through first-order parameter updates, yet it remains unclear whether such updates preserve logical coherence. We invest

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

DynaPURLS: Dynamic Refinement of Part-Aware Representations for Skeleton-Based Zero-Shot Action Recognition

DGX agent

arXiv:2512.11941v2 Announce Type: replace-cross Abstract: Zero-shot skeleton-based action recognition (ZS-SAR) is fundamentally constrained by prevailing approaches that rely on aligning skeleton feat

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

E = T*H/(O+B): A Dimensionless Control Parameter for Mixture-of-Experts Ecology

DGX agent

arXiv:2605.06415v2 Announce Type: replace-cross Abstract: We introduce E = T*H/(O+B), a dimensionless control parameter that predicts whether Mixture-of-Experts (MoE) models will develop a healthy exp

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

EchoDistill:Alignment Noisy-to-Clean Self-Distillation for Robust Audio LLMs

DGX agent

arXiv:2605.23954v1 Announce Type: cross Abstract: Audio Large Language Models (ALLMs) are highly vulnerable to real-world noise, which often induces severe semantic drift and hallucinations. Existing

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

EchoPilot: Training-Free Ultrasound Video Segmentation via Scale-Space Semantic Prompting and Reliability-Gated Memory

DGX agent

arXiv:2605.25944v1 Announce Type: cross Abstract: Ultrasound video segmentation is clinically valuable yet difficult due to speckle noise, weak boundaries, and rapid anatomical deformation. Recent pro

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Efficient Benchmarking Is Just Feature Selection and Multiple Regression

DGX agent

arXiv:2605.25773v1 Announce Type: cross Abstract: Efficient benchmarking techniques aim to lower the computational cost of evaluating LLMs by predicting full benchmark scores using only a subset of a

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Efficient DP-SGD for LLMs with Randomized Clipping

DGX agent

arXiv:2605.24879v1 Announce Type: new Abstract: Large language models (LLMs) are trained on vast datasets that may contain sensitive information. Differential privacy (DP), the de facto standard for f

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Efficient Long-Horizon Vision-Language-Action Models via Static-Dynamic Disentanglement

DGX agent

arXiv:2602.03983v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for generalist robotic control. Built upon vision-language model (

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

EMA-Nesterov: Stabilizing Nesterov's Lookahead for Accelerated Deep Learning Optimization

DGX agent

arXiv:2605.25395v1 Announce Type: new Abstract: Lookahead-based acceleration methods, such as Nesterov's momentum, are widely used in optimization, but they often become unreliable in deep learning tr

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Emission-Aware Reinforcement Learning for Sustainable Electric Vehicle Charging and Carbon Dioxide Reduction Under Varying Renewable Penetration

DGX agent

arXiv:2605.24543v1 Announce Type: new Abstract: The rapid growth of Electric Vehicle (EV) adoption challenges power distribution networks through peak load spikes, voltage instability, and transformer

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Emotional intelligence in large language models is fragmented across perception, cognition, and interaction

DGX agent

arXiv:2605.24686v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly integrated into emotionally sensitive domains, the structural integrity of their emotional intelligence

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Empirical Analysis and Detection of Hallucinations in LLM-Generated Bug Report Summaries

DGX agent

arXiv:2605.24137v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to generate summaries of software bug reports, including sections such as Steps-to-Reproduce (S2R),

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Enhancing Reliability in LLM-Based Secure Code Generation

DGX agent

arXiv:2605.24300v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for code generation, but their security reliability remains inconsistent across languages and prompting s

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Equation-Free Coarse Control of Distributed Parameter Systems via Local Neural Operators

DGX agent

arXiv:2509.23975v2 Announce Type: replace-cross Abstract: The control of high-dimensional distributed parameter systems (DPS) remains a challenge when explicit coarse-grained equations are unavailable

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

ERNIE-Image Technical Report

DGX agent

arXiv:2605.25347v1 Announce Type: cross Abstract: We introduce ERNIE-Image, an open-source text-to-image generation model built upon an 8B single-stream DiT architecture. ERNIE-Image aims to bridge th

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

everybody talks about the china->us catchup not enough people talking about the us-> china catchup great job @o_lacombe et al, @robert_mchar…

DGX agent

everybody talks about the china->us catchup not enough people talking about the us-> china catchup great job @o_lacombe et al, @robert_mchardy et al! [AINews 3 Apr 2026] Gemma 4: The world's best smal

model-releasesswyx--x
26 May 2026
Model Releases

EvoCode-Bench: Evaluating Coding Agents in Multi-Turn Iterative Interactions

DGX agent

arXiv:2605.24110v1 Announce Type: new Abstract: Coding agents are increasingly used as iterative development partners, but most benchmarks still evaluate one specification followed by one final assess

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

EvoEGF-Mol: Evolving Exponential Geodesic Flow for Structure-based Drug Design

DGX agent

arXiv:2601.22466v2 Announce Type: replace Abstract: Structure-Based Drug Design (SBDD) aims to discover bioactive ligands. Conventional approaches construct probability paths separately in Euclidean a

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Exploration of Perceptual Speech Features for Clinical Decision-Support in Mental Health Care

DGX agent

arXiv:2605.24678v1 Announce Type: new Abstract: Speech and language technologies offer valuable opportunities for supporting mental health assessment through objective and interpretable cues. We prese

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Explore Before You Solve: The Speed--Depth Trade-off in Epistemic Agents for ARC-AGI-3

DGX agent

arXiv:2605.25931v1 Announce Type: new Abstract: We systematically investigate all 25 public ARC-AGI-3 games and find that every one is reachable through non-intelligent strategies: 10 in a single blin

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Extending Embodied Question Answering from Perception to Decision

DGX agent

arXiv:2605.25813v1 Announce Type: new Abstract: Embodied Question Answering (EQA) connects perception, reasoning, and interaction within embodied environments. However, existing datasets and benchmark

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth

DGX agent

arXiv:2605.25052v1 Announce Type: new Abstract: Chains of thought (CoTs) have become central in interpreting and auditing behaviors of large language models. Yet growing evidence suggests that these t

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Feature Learning in Wide Neural Networks under muP: Identifiability and Sparse-Dictionary Decomposition of the Mean-Field Limit

DGX agent

arXiv:2605.24710v1 Announce Type: new Abstract: We establish four structural results for feature learning in wide two-layer neural networks under the Maximal Update Parametrization (muP). First, we pr

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model

DGX agent

arXiv:2510.10921v3 Announce Type: replace-cross Abstract: Fine-grained vision-language understanding requires precise alignment between visual content and linguistic descriptions, a capability that re

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Fine-Tuning and Serving Gemma 4 31B on Google Cloud TPU: A Technical Comparison with GPU Baselines

DGX agent

arXiv:2605.25645v1 Announce Type: cross Abstract: We present the first end-to-end demonstration of fine-tuning and serving Google's Gemma 4 31B model on TPU hardware, providing an empirical comparison

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Fine-Tuning Language Models to Know What They Know

DGX agent

arXiv:2602.02605v2 Announce Type: replace-cross Abstract: Evaluating true metacognition in Large Language Models (LLMs) is difficult due to biases and heuristics. This paper presents a framework to me

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

FLOATBench: A Dataset and Benchmark for Floating Offshore Wind Turbine Tower Fatigue

DGX agent

arXiv:2605.25717v1 Announce Type: new Abstract: Most of the world's offshore wind resource lies in waters too deep for fixed-bottom foundations, making floating offshore wind turbines (FOWTs) essentia

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

FloorplanQA: A Benchmark for Spatial Reasoning in LLMs using Structured Representations

DGX agent

arXiv:2507.07644v4 Announce Type: replace Abstract: We introduce FloorplanQA, a diagnostic benchmark for evaluating spatial reasoning in large language models (LLMs). FloorplanQA is grounded in struct

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

FLoRIST: Singular Value Thresholding for Efficient and Accurate Federated Fine-Tuning of Large Language Models

DGX agent

arXiv:2506.09199v2 Announce Type: replace-cross Abstract: Integrating Low-Rank Adaptation (LoRA) into federated learning offers a promising solution for parameter-efficient fine-tuning of Large Langua

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

FoodMonitor: Benchmarking MLLMs for Explainable Compliance Analysis

DGX agent

arXiv:2605.24503v1 Announce Type: cross Abstract: As AI-powered compliance monitoring becomes increasingly important in public governance and industrial safety, the ability to provide verifiable evide

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Found in Conversation: LLMs Teach Themselves to Close the Multi-Turn Gap

DGX agent

arXiv:2605.24432v1 Announce Type: new Abstract: Large Language Model (LLM) interactions are typically underspecified, with users clarifying all necessary details across multiple conversational turns.

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

FOUND-IT: Foundation-model-first Task-driven 3D Scene Graphs with Granularity on Demand

DGX agent

arXiv:2605.25371v1 Announce Type: new Abstract: We present the first approach to build hierarchical task-driven 3D scene graphs of arbitrary indoor or outdoor environments using an uncalibrated monocu

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Fourier Feature Pyramids for Physics-Informed Neural Networks

DGX agent

arXiv:2605.24278v1 Announce Type: new Abstract: We present an improved neural field architecture for solving partial differential equations (PDEs). Current physics-informed neural networks (PINNs) pro

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Free the 100B Gemma 4 MoE! Gemini Flash 3.5 is out so now you can release it!

DGX agent

Clem Delangue advocates for the release of a 100 billion parameter Gemma 4 Mixture of Experts model, suggesting that Gemini Flash 3.5's release creates an opportunity for this larger model to be made

model-releasesclem-delangue--x
26 May 2026
Model Releases

From DPPs to k-DPPs: identifiability analysis via spectral decomposition

DGX agent

arXiv:2605.25526v1 Announce Type: cross Abstract: We study the geometry of determinantal point processes (DPPs) through the spectral decomposition L=ULambda U^{op}. The spectrum Lambda governs the car

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

From Facts to Insights: A Persona-Driven Dual Memory Framework and Dataset for Role-Playing Agents

DGX agent

arXiv:2605.25693v1 Announce Type: new Abstract: While role-playing agents excel in short-term interactions, long-term conversations overwhelm context windows, motivating external memory frameworks. Cu

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

From Index to Equity: Pre-Training Transformers for Stock Return Prediction

DGX agent

arXiv:2605.23962v1 Announce Type: cross Abstract: This research aims to leverage machine learning to improve stock price prediction and support informed investment decisions related to buying, selling

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

DGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

From One-Pass SGD to Data Reuse: Mini-Batch Scaling Laws in Sketched Linear Regression

DGX agent

arXiv:2605.24316v1 Announce Type: new Abstract: Scaling laws provide compact descriptions of how prediction error varies with compute, model size, and data, but existing theory mainly treats single-sa

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

From Prompt Optimization to Multi-Dimensional Credibility Evaluation: Enhancing Trustworthiness of Chinese LLM-Generated Liver MRI Reports -- with Preliminary Extension to Lung Cancer

DGX agent

arXiv:2510.23008v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated promising performance in generating diagnostic conclusions from imaging findings, thereby supporting

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

From Theory to Decision Rule: Calibrating the Noisy-Label Crossover for Vision-Language Model Weak Supervision Across Three Medical-Imaging Benchmarks

DGX agent

arXiv:2605.24771v1 Announce Type: cross Abstract: Classical noisy-label theory predicts that downstream performance under weak supervision is bounded above by the labeler's accuracy, implying a sharp

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

FrontierOR: Benchmarking LLMs' Capacity for Efficient Algorithm Design in Large-Scale Optimization

DGX agent

arXiv:2605.25246v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for optimization modeling and solver-code generation, yet practical operations research and optimizat

model-releasesarxiv-cs-ai
26 May 2026
← Previous
1…264265266267268…472
Next →