AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,620 results
29 May 2026

CosmicFish-HRM: Adaptive Reasoning via Hierarchical Recurrent Mechanisms in Compact Language Models

Model ReleasesDGX agent

arXiv:2605.28919v1 Announce Type: cross Abstract: Large language models have achieved strong reasoning capabilities, though often at the cost of massive parameter counts and expensive inference. In th

CriticalKV: Optimizing KV Cache Eviction from an Output Perturbation Perspective

Model ReleasesDGX agent

arXiv:2502.03805v2 Announce Type: replace Abstract: Large language models have revolutionized natural language processing but face significant challenges of high storage and runtime costs, due to the

CrystalXRD-Bench: Benchmarking Vision-Language Models for XRD Peak Indexing Across Diverse Crystalline Materials

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.29446v1 Announce Type: new Abstract: Miller-index identification from powder XRD patterns requires capabilities untested by existing multimodal benchmarks: the model must read a narrow peak

Deep Adaptive Dimension Reduction for Bayesian Inference in Inverse Problems

Model ReleasesDGX agent

arXiv:2605.29373v1 Announce Type: new Abstract: Solving high-dimensional PDE-governed inverse problems is often challenging due to complex non-Gaussian posterior distributions, expensive forward model

Demystifying Scientific Problem-Solving in LLMs by Probing Knowledge and Reasoning

Model ReleasesDGX agent

arXiv:2508.19202v3 Announce Type: replace Abstract: Scientific problem solving poses unique challenges for LLMs, requiring both deep domain knowledge and the ability to apply such knowledge through co

DenseSteer: Steering Small Language Models towards Dense Math Reasoning

Model ReleasesDGX agent

arXiv:2605.29247v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong chain-of-thought (CoT) reasoning abilities, while smaller models (<= 3B parameters) significantly underp

Developer's guide to Gemini Enterprise and A2UI integration

Model ReleasesDGX agent

If you've built a chatbot, you know this conversation: User: 'Book a table for two tomorrow at 7pm.' Agent: 'Okay, for what day?' User: 'Tomorrow.' Agent: 'What time?' A date picker would have ended t

Dial HEALTHDIAL for Advice: A Multilingual and Multi-Parallel Spoken Dialogue Dataset for Knowledge-Grounded Information Seeking

Model ReleasesDGX agent

arXiv:2605.30107v1 Announce Type: new Abstract: Creating spoken dialogue datasets is methodologically challenging, and these challenges are amplified when the goal is to build multilingual, multi-para

Differentiable Belief-based Opponent Shaping

Model ReleasesDGX agent

arXiv:2605.29042v1 Announce Type: new Abstract: Human coordination often relies on the ability to influence the beliefs of others through strategic action. In multi-agent reinforcement learning, oppon

DiffSpot: Can VLMs Spot Fine-Grained Visual Differences in Web Interfaces?

Model ReleasesDGX agent

arXiv:2605.29615v1 Announce Type: cross Abstract: Vision-language models (VLMs) have made strong progress on high-level image-text alignment, yet their ability to perceive subtle visual differences re

Diffusion-based learning framework for Constrained Nonconvex Optimization with Weighted Bootstrapped Refinement

Model ReleasesDGX agent

arXiv:2502.10330v4 Announce Type: replace Abstract: Recent advances in diffusion models show promising potential to accelerate nonconvex problem solving by leveraging their multimodality. However, mos

Diffusion differentiable resampling

Model ReleasesDGX agent

arXiv:2512.10401v3 Announce Type: replace-cross Abstract: This paper is concerned with differentiable resampling in the context of sequential Monte Carlo (e.g., particle filtering). Drawing on reparam

DirectorBench: Diagnosing Long-Form Video Generation with Personalized Multi-Agent Evaluation

Model ReleasesDGX agent

arXiv:2605.30090v1 Announce Type: new Abstract: Long-form video generation is rapidly moving from short, single-scene synthesis toward minute-long, multi-shot creation with narrative structure, cinema

Dissecting the Black Box: Circuit-Level Analysis of LLM Vulnerability Detection

Model ReleasesDGX agent

arXiv:2605.29901v1 Announce Type: cross Abstract: Large language models (LLMs) can detect software vulnerabilities, but how do they actually identify vulnerable code? We address this question using me

DMC-CF: Dynamic Multimodal CounterFactual QA benchmark for Causal Reasoning

Model ReleasesDGX agent

arXiv:2605.29339v1 Announce Type: new Abstract: With the rapid advancement of multimodal large language models (MLLMs), models have demonstrated increasingly powerful multimodal capabilities. However,

Do Physics Foundation Models Learn Generalizable Physics? A Bias-Aware Benchmark Across Physical Regimes and Distribution Shifts

Model ReleasesDGX agent

arXiv:2605.29283v1 Announce Type: cross Abstract: Recent physics foundation models claim general spatiotemporal forecasting ability, yet their evaluations often collapse performance into a single aver

DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark

Model ReleasesDGX agent

arXiv:2605.30027v1 Announce Type: new Abstract: Multimodal documents contain diverse elements, such as tables, figures, and layouts, which can complicate retrieval tasks. While current approaches typi

Document Parsing + Gemini 🔥 Excited to collaborate with the Google team on this, here's to many more!

Model ReleasesDGX agent

Document Parsing + Gemini 🔥 Excited to collaborate with the Google team on this, here's to many more! The team at @llama_index built an awesome template using LlamaParse and the new Managed Agents in

Draw a path that a drone can follow https://x.com/bilawalsidhu/status/2059419767417487718?s=20

Model ReleasesDGX agent

This post likely describes a tool or system for planning autonomous drone flight paths, possibly involving visualization or interactive path-drawing interfaces. It may demonstrate how users can create

Dynamic Mixture of Progressive Parameter-Efficient Expert Library for Lifelong Robot Learning

Model ReleasesDGX agent

arXiv:2506.05985v3 Announce Type: replace Abstract: A generalist agent must continuously learn and adapt throughout its lifetime, achieving efficient forward transfer while minimizing catastrophic for

Dynamics of Stochastic Momentum with Sparse Updates in High Dimensions

Model ReleasesDGX agent

arXiv:2605.28961v1 Announce Type: cross Abstract: Existing theory of momentum assumes that gradients arrive at every parameter at a roughly constant rate, an assumption violated in practice by heavy-t

DynSess: Dynamic Session-Level Evaluation and Optimization Framework for Role-Playing Agents

Model ReleasesDGX agent

arXiv:2605.29256v1 Announce Type: cross Abstract: Role-playing with large language models is fundamentally a session-level task, requiring agents to sustain character identity and interaction quality

DySem: Uncovering Dynamic Semantic Components via Multilingual Consensus for Calculating Semantic Textual Similarity

Model ReleasesDGX agent

arXiv:2605.29751v1 Announce Type: new Abstract: Calculating semantic textual similarity is a foundational task in natural language processing. Current large language models (LLMs) based methods typica

EarthShift: a benchmark for measuring robustness to real-world distribution shifts in Earth observation

Model ReleasesDGX agent

arXiv:2605.29330v1 Announce Type: new Abstract: Current Earth observation benchmarks focus on measuring performance on diverse tasks and applications, typically measuring generalization in-distributio

ElevenLabs launches Dubbing v2, which it says preserves the original speaker's emotion, tone, and pacing across 90+ languages while staying synced to content (ElevenLabs)

Model ReleasesDGX agent

ElevenLabs: ElevenLabs launches Dubbing v2, which it says preserves the original speaker's emotion, tone, and pacing across 90+ languages while staying synced to content — Today we're launching Dubbin

Embodied3DBench: Benchmarking Low-Level Embodied Spatial Intelligence of Vision Language Models

Model ReleasesDGX agent

arXiv:2605.29074v1 Announce Type: new Abstract: Are current Vision Language Models (VLMs) ready to comprehend and reason about complex embodied interactions in 3D environments? We introduce Embodied3D

Empathic Prompting: Non-Verbal Context Integration for Multimodal LLM Conversations

Model ReleasesDGX agent

arXiv:2510.20743v2 Announce Type: replace-cross Abstract: We present Empathic Prompting, a novel framework for multimodal human-AI interaction that enriches Large Language Model (LLM) conversations wi

Entity-Collision: A Stratified Protocol for Attributing Retrieval Lift in Agent Memory

Model ReleasesDGX agent

arXiv:2605.29630v1 Announce Type: cross Abstract: End-to-end agent-memory benchmarks report a single hit@k per retriever, confounding lexical leakage (uncontrolled query/gold/distractor entity overlap

ESPO: Early-Stopping Proximal Policy Optimization

Model ReleasesDGX agent

arXiv:2605.29860v1 Announce Type: cross Abstract: When a large language model under reinforcement learning commits a wrong reasoning step early in a trajectory, standard algorithms force it to keep ge

Evaluating Cross-lingual Knowledge Consistency in Code-Mixed vis-a-vis Indian Languages using IndicKLAR

Model ReleasesDGX agent

arXiv:2605.29637v1 Announce Type: new Abstract: Large language models recall knowledge reliably in English but often fail on the same query posed in a lower-resourced language -- a crosslingual consis

Evaluating Dataset Watermarking for Fine-tuning Traceability of Customized Diffusion Models: A Comprehensive Benchmark and Removal Approach

Model ReleasesDGX agent

arXiv:2511.19316v2 Announce Type: replace-cross Abstract: Recent fine-tuning techniques for diffusion models enable them to reproduce specific image sets, such as particular faces or artistic styles,

EVL-ECG: Efficient ECG Interpretation With Multi-Aspect Heterogeneous Knowledge Distillation

Model ReleasesDGX agent

arXiv:2605.29977v1 Announce Type: new Abstract: High-fidelity ECG interpretation is increasingly reliant on massive foundation models, yet their deployment in clinical edge-care remains hindered by ex

Evolutionary Dynamics of Cooperation in Next-Generation LLM Agent Systems: A Cross-Provider Empirical Extension

Model ReleasesDGX agent

arXiv:2605.29874v1 Announce Type: cross Abstract: Do next-generation LLM agents inherit the cooperative biases documented in their predecessors, or does scale and provider diversity reshape equilibriu

Evolutionary Rule Extraction from Corporate Default Prediction Models

Model ReleasesDGX agent

arXiv:2605.29478v1 Announce Type: cross Abstract: Small and medium-sized enterprises (SMEs) represent the majority of firms in most economies and often face financial constraints and higher vulnerabil

ExCAM: Explainable Cultural Awareness Metrics

Model ReleasesDGX agent

arXiv:2605.29897v1 Announce Type: new Abstract: Evaluating the cultural awareness of large language models is crucial to ensure the fairness of generated text and the generalizability of applications

Fairness-Aware Federated Learning with Trajectory Shapley Value

Model ReleasesDGX agent

arXiv:2605.30336v1 Announce Type: new Abstract: Federated learning is an emerging distributed paradigm that addresses the challenges posed by heterogeneous, privacy-sensitive data. It enables multiple

Fairness Beyond Demographics: Optimizing Performance Across Appearance-Based Hidden Cohorts in Medical Imaging

Model ReleasesDGX agent

arXiv:2605.29827v1 Announce Type: new Abstract: Medical image analysis models can exhibit performance disparities across patient subgroups, threatening clinical safety and fairness. Existing methods t

FarSkip-Collective: Unhobbling Blocking Communication in Mixture of Experts Models

Model ReleasesDGX agent

arXiv:2511.11505v3 Announce Type: replace Abstract: Blocking communication presents a major hurdle in running MoEs efficiently in distributed settings. To address this, we present FarSkip-Collective w

Feature Geometry of LoRA Adapters: A Sparse Autoencoder Analysis of Representational Divergence in Fine-Tuned Language Models

Model ReleasesDGX agent

arXiv:2605.28896v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has emerged as a widely adopted approach for adapting large language models, yet the internal representational changes induce

FedQHD: Closed-Form Function-Space Federated Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.29002v1 Announce Type: new Abstract: Federated reinforcement learning enables decentralized agents to collaboratively improve policies or value estimates without exchanging raw trajectories

Fine-tune your first AI model today. Run GPT4o level model and run on your phone or laptop. @OpenBMB released 15M samples SFT dataset that y…

Model ReleasesDGX agent

Fine-tune your first AI model today. Run GPT4o level model and run on your phone or laptop. @OpenBMB released 15M samples SFT dataset that you can use right now. (319GB high-quality post training data

FinGuard: Detecting Financial Regulatory Non-Compliance in LLM Interactions

Model ReleasesDGX agent

arXiv:2605.29427v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed in financial services, a single non-compliant interaction can expose institutions to regulator

FinVerBench: Benchmark Validity and Calibration in Large Language Model Financial Statement Verification

Model ReleasesDGX agent

arXiv:2605.29586v1 Announce Type: new Abstract: We introduce FinVerBench, a benchmark and validity study for financial statement verification: determining whether a set of corporate financial statemen

First head-to-head comparison of agentic AI applied to the analysis of simulated data of the Einstein Telescope

Model ReleasesDGX agent

arXiv:2605.28916v1 Announce Type: cross Abstract: We report a comparison of two state-of-the-art agentic AI systems, Claude Code (Anthropic) and Codex (OpenAI), tasked with autonomously executing a si

FoRA: Fisher-orthogonal Rank Adaptation for Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2605.29317v1 Announce Type: new Abstract: Parameter-efficient fine-tuning(PEFT) has largely focused on LoRA and its accuracy-oriented variants, leaving the original goal of reducing trainable pa

FormInv: A Measurement Protocol for Semantic Invariance in Mathematical Reasoning Benchmarks

Model ReleasesDGX agent

arXiv:2605.29001v1 Announce Type: cross Abstract: A paraphrase-quality audit of MathCheck (ICLR 2025) detected 4 semantically incorrect paraphrases in 129 groups (3.1%); removing them drops GPT-4o fro

FPLIER: Federated Pathway-Level Information Extractor

Model ReleasesDGX agent

arXiv:2605.29587v1 Announce Type: cross Abstract: In transcriptomics, gene-set-aware factorization methods such as the Pathway Level Information Extractor (PLIER) are most effective when trained on la

From Sublinear to Linear: Local Convergence in Finite-Width Networks via Locally Polyak-Lojasiewicz Regions

Model ReleasesDGX agent

arXiv:2507.21429v3 Announce Type: replace-cross Abstract: We study local linear convergence of gradient descent for finite-width feedforward networks under the squared empirical loss. Prior work shows

From XXLTraffic to EvoXXLTraffic: Scaling Traffic Forecasting to Sensor-Evolving Networks

Model ReleasesDGX agent

arXiv:2605.29768v1 Announce Type: new Abstract: Existing traffic forecasting benchmarks assume a fixed sensor set, but real road-sensor networks grow continuously as the road network changes year by y

Frontier LLM-based agents can overcome the ontology curation bottleneck for natural phenotypes

Model ReleasesDGX agent

arXiv:2605.28965v1 Announce Type: new Abstract: Linking free-text phenotype descriptions to ontology terms, typically referred to as phenotype annotation, is essential for the cross-study integration

GenEraser: Generalizable Video Object Removal via Balanced Text-Mask Guidance and Decoupled Locator-Preserver

Model ReleasesDGX agent

arXiv:2605.30045v1 Announce Type: new Abstract: Video object removal frequently struggles to simultaneously eliminate target objects and their associated physical effects (e.g., smoke, reflections, li

GEO-Bench: Benchmarking Ranking Manipulation in Generative Engine Optimization

Model ReleasesDGX agent

arXiv:2605.29107v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly rank products, documents, and recommendations for user queries, which makes manipulating these rankings a gr

Give it Space! Explicit Disentangling of Positional and Semantic Representations in Encoders

Model ReleasesDGX agent

arXiv:2605.30022v1 Announce Type: cross Abstract: Positional encoding (PE) underpins how permutation-invariant Transformers represent sequence order, yet how positional information is processed and st

Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning

Model ReleasesDGX agent

arXiv:2602.01058v2 Announce Type: replace-cross Abstract: Post-training of reasoning LLMs is a holistic process that typically consists of an offline SFT stage followed by an online reinforcement lear

GPF-LiveNews: A Streaming Evaluation Protocol for Group-Conditioned Framing in Large Language Models

Model ReleasesDGX agent

arXiv:2605.28848v1 Announce Type: cross Abstract: Deployed language models are evaluated in a non-stationary environment: model versions, retrieval layers, safety systems, and real-world inputs all ch

GPIC: A Giant Permissive Image Corpus for Visual Generation

Model ReleasesDGX agent

arXiv:2605.30341v1 Announce Type: cross Abstract: Studying scalable methods for visual generative modeling requires large, accessible, and stable datasets. We introduce GPIC, a Giant Permissive Image

GPT-5.5 Pro is the model producing many of the novel math proofs, and also the model you should have reviewing any technical or academic pap…

Model ReleasesDGX agent

I cannot verify this claim as it references a future date (the URL timestamp appears invalid) and describes a model version (GPT-5.5 Pro) that doesn't exist in current public releases. The post appear

Gradient Perturbation: Learning to Perturb Gradients for Adaptive Training

Model ReleasesDGX agent

arXiv:2605.29494v1 Announce Type: new Abstract: Deep neural network training involves both forward propagation (from features through logits to loss) and backward propagation (from loss through gradie

Gram: Assessing sabotage propensities via automated alignment auditing

Model ReleasesDGX agent

arXiv:2605.30322v1 Announce Type: cross Abstract: We introduce Gram, an automated alignment auditing framework to assess the propensity of AI agents to engage in sabotage. We evaluate Gemini models ac

GRASP: Gated Regression-Aware Skill Proposer for Self-Improving LLM Agents

Model ReleasesDGX agent

arXiv:2605.29668v1 Announce Type: new Abstract: LLM agents acting in structured environments fail in operational rather than conversational ways, and reliability depends on procedural knowledge of the

← Previous
1…193194195196197…377
Next →