AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
All
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,628 results
Model Releases

Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model

DGX agent

arXiv:2412.07333v2 Announce Type: replace-cross Abstract: Pose-Guided Person Image Synthesis (PGPIS) aims to generate human images in specified poses while preserving the identity and appearance of a

model-releasesarxiv-cs-ai
26 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Game-Theoretic Modeling of Heterogeneous Investor Interactions for Stock Price Forecasting

DGX agent

arXiv:2605.23953v1 Announce Type: cross Abstract: Accurate stock price forecasting has consistently remained a pivotal yet challenging FinTech task that underpins quantitative trading and investment d

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

GDformer: Going Beyond Subsequence Isolation for Multivariate Time Series Anomaly Detection

DGX agent

arXiv:2501.18196v3 Announce Type: replace Abstract: Unsupervised anomaly detection of multivariate time series is a challenging task, given the requirements of deriving a compact detection criterion w

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open …

DGX agent

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open models. Some ideas for what comes next, May 2026 Gemini Flas

model-releasesclem-delangue--x
26 May 2026
Model Releases

Generalizable Vision-Language Few-Shot Adaptation with Predictive Prompts and Negative Learning

DGX agent

arXiv:2505.11758v2 Announce Type: replace-cross Abstract: Few-shot adaptation of vision-language models remains fundamentally limited by how negative class signals are handled at inference. Existing m

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Geo-Expert: Towards Expert-Level Geological Reasoning via Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2605.24844v1 Announce Type: new Abstract: While general-purpose Large Language Models (LLMs) applied to Geology often hallucinate when reasoning about subsurface structures and deep-time evoluti

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

GL-LFGNN:A Global-Local Dual-branch Causal Graph Neural Network Based on Liang-Kleeman Information Flow for EEG Emotion Recognition

DGX agent

arXiv:2605.25061v1 Announce Type: cross Abstract: EEG-based emotion recognition holds significant promise for objective diagnosis of mood disorders. Graph neural networks (GNNs) have emerged as the do

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

GlobalDentBench: A Multinational Benchmark for Evaluating LLM Clinical Reasoning in Dentistry with Expert Calibration

DGX agent

arXiv:2605.24636v1 Announce Type: new Abstract: While large language models (LLMs) hold transformative potential for medicine, their reasoning robustness and safety in real-world clinical scenarios re

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Goal-driven Bayesian Optimal Experimental Design for Robust Decision-Making Under Model Uncertainty

DGX agent

arXiv:2605.26093v1 Announce Type: new Abstract: Bayesian optimal experimental design (BOED) selects experiments to maximize information gain about model parameters. However, in decision-critical setti

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers

DGX agent

arXiv:2605.24518v1 Announce Type: cross Abstract: The quadratic complexity of self-attention in Transformer models remains a significant bottleneck for processing long sequences and deploying large la

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

GreenSeg: Ground Segmentation Algorithm for Agricultural Robots in Mediterranean Greenhouses using RGB-D Point Clouds

DGX agent

arXiv:2605.25279v1 Announce Type: new Abstract: Greenhouse agriculture in the Mediterranean region faces significant automation challenges due to its unique structural and environmental constraints. T

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

GroupTravelBench: Benchmarking LLM Agents on Multi-Person Travel Planning

DGX agent

arXiv:2605.25200v1 Announce Type: new Abstract: Travel planning is a realistic task for evaluating the planning and tool-use abilities of LLM agents. However, existing benchmarks typically assume only

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Guided Flow Matching for Forward and Inverse PDE Problems with Sparse Observations: Algorithm and Theory

DGX agent

arXiv:2605.25509v1 Announce Type: cross Abstract: Reconstructing PDE solutions from sparse observations is a core challenge in scientific computing. We present FM4PDE, a flow-matching generative frame

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Hadamard Representation: Scaffolding Performance Across Model-free RL

DGX agent

arXiv:2406.09079v5 Announce Type: replace Abstract: Deep reinforcement learning agents progressively lose representational capacity during training: neurons become dormant, removing active capacity fr

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

HEAPr: Hessian-based Efficient Atomic Expert Pruning in Output Space

DGX agent

arXiv:2509.22299v3 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) architectures in large language models (LLMs) deliver exceptional performance and reduced inference costs compared to

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

HiGraph: A Large-Scale Hierarchical Graph Dataset for Malware Analysis

DGX agent

arXiv:2509.02113v2 Announce Type: replace-cross Abstract: The advancement of graph-based malware analysis is critically limited by the absence of large-scale datasets that capture the inherent hierarc

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

HiMed: Incentivizing Hindi Reasoning in Medical LLMs

DGX agent

arXiv:2605.24635v1 Announce Type: new Abstract: Medical large language models hold promise for reducing healthcare disparities, yet Hindi remains severely underrepresented. While medical LLMs excel in

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing

DGX agent

arXiv:2605.24687v1 Announce Type: cross Abstract: Text-to-Image (T2I) models have made significant strides in visual realism and semantic consistency, yet they often perpetuate and amplify societal bi

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

How Many Tools Should an LLM Agent See? A Chance-Corrected Answer

DGX agent

arXiv:2605.24660v1 Announce Type: cross Abstract: Before an LLM agent can use a tool, a retrieval system must decide which candidate tools to show to the agent. How long should that shortlist be? Show

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

How Much Do Large Language Model Cheat on Evaluation? Benchmarking Overestimation under the One-Time-Pad-Based Framework

DGX agent

arXiv:2507.19219v2 Announce Type: replace Abstract: Overestimation in evaluating large language models (LLMs) has become an increasing concern. Due to the contamination of public benchmarks or imbalan

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

How Much Thinking is Enough? Quantifying and Understanding Redundancy in LLM Reasoning

DGX agent

arXiv:2605.23926v1 Announce Type: new Abstract: Reasoning-capable large language models solve hard problems by emitting long chains of thought, paying heavily in latency, GPU time, and energy. Casual

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

How we evolved Google’s global and data center networks for the AI era

DGX agent

Over the last 25 years of building Google’s global network, we’ve navigated major architectural eras — from the Internet, to streaming, and the cloud. Today, we are squarely in the midst of a fourth:

model-releasesgoogle-cloud-ai
26 May 2026
Model Releases

How Well Do Models Follow Their Constitutions?

DGX agent

arXiv:2605.24229v1 Announce Type: new Abstract: Frontier AI developers now train models against long written behavioral specifications, such as Anthropic's constitution (Anthropic, 2025a) and OpenAI's

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Hylos: Operability Contracts for Model-Native Spatial Intelligence

DGX agent

arXiv:2605.24728v1 Announce Type: new Abstract: Foundation models can increasingly describe, reconstruct, and generate 3D objects, assemblies, scenes, and environments, but visually plausible spatial

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

I don't comment on every article that uses out-of-date measures on AI ability, but I felt (probably wrongly) that the article was a response…

DGX agent

I don't comment on every article that uses out-of-date measures on AI ability, but I felt (probably wrongly) that the article was a response to my viral tweet, so I felt I needed to say something! htt

model-releasesethan-mollick--x
26 May 2026
Model Releases

I found this Wired article on AI fact-checking frustrating. It could have been about why we continue to need human fact checkers (talk to pe…

DGX agent

I found this Wired article on AI fact-checking frustrating. It could have been about why we continue to need human fact checkers (talk to people, use judgement, resolve conflict). Instead it is full o

model-releasesethan-mollick--x
26 May 2026
Model Releases

IndexMem: Learned KV-Cache Eviction with Latent Memory for Long-Context LLM Inference

DGX agent

arXiv:2605.25475v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly expected to operate over long contexts, yet standard softmax attention incurs a KV cache that grows line

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

INDUCTION: Finite-Structure Concept Synthesis in First-Order Logic

DGX agent

arXiv:2602.18956v3 Announce Type: replace Abstract: We introduce INDUCTION, a benchmark for finite structure concept synthesis in first order logic. Given small finite relational worlds with extension

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Inference Time Optimization with Confidence Dynamics

DGX agent

arXiv:2605.25244v1 Announce Type: new Abstract: Inference time optimization techniques, such as repeated sampling, have significantly advanced the reasoning capabilities of Large Language Models (LLMs

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

InfiFPO: Implicit Model Fusion via Preference Optimization in Large Language Models

DGX agent

arXiv:2505.13878v3 Announce Type: replace-cross Abstract: Model fusion combines multiple Large Language Models (LLMs) with different strengths into a more powerful, integrated model through lightweigh

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Insuring Every Action: An Authority Frontier Framework for Runtime Actuarial Control of Autonomous AI Agents

DGX agent

arXiv:2605.25632v1 Announce Type: new Abstract: Autonomous AI agents increasingly issue side-effect-bearing actions: database mutations, refunds, payments, external commitments. We propose the Actuari

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 …

DGX agent

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 apps + 200+ MCP tools + 1,290 skills + process / outcome rew

model-releasesclem-delangue--x
26 May 2026
Model Releases

IterInject: Indirect Prompt Injection Against LLM Agents via Feedback-Guided Iterative Optimization

DGX agent

arXiv:2605.24659v1 Announce Type: new Abstract: LLM-based agents are increasingly deployed for complex tasks requiring planning, tool use, and interaction with external services. Their reliance on unt

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

JacQuant: STE-Free Quantization-Aware Training via Learned Jacobian Surrogates

DGX agent

arXiv:2605.25469v1 Announce Type: new Abstract: Quantization-aware training (QAT) is widely deployed but typically relies on the Straight-Through Estimator (STE), which passes gradients through non-di

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

JAEGER: Joint 3D Audio-Visual Grounding and Reasoning in Simulated Physical Environments

DGX agent

arXiv:2602.18527v2 Announce Type: replace-cross Abstract: Current audio-visual large language models (AV-LLMs) are predominantly restricted to 2D perception, relying on RGB video and monaural audio. T

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

JEPA-DNA: Grounding Genomic Foundation Models through Joint-Embedding Predictive Architectures

DGX agent

arXiv:2602.17162v2 Announce Type: replace Abstract: Genomic Foundation Models (GFMs) typically rely on Masked Language Modeling (MLM) or Next-Token Prediction (NTP) to learn the 'Laws of Nature'. Whil

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

JudgmentBench: Comparing Rubric and Preference Evaluation for Quality Assessment

DGX agent

arXiv:2605.25240v1 Announce Type: cross Abstract: Two methodologies dominate current practices of benchmarking: rubric-based scoring evaluates items against predefined criteria, whereas comparative ju

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

KAME: Tandem Architecture for Enhancing Knowledge in Real-Time Speech-to-Speech Conversational AI

DGX agent

arXiv:2510.02327v2 Announce Type: replace-cross Abstract: Real-time speech-to-speech (S2S) models excel at generating natural, low-latency conversational responses but often lack deep knowledge and se

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Keep the Proof State Live: Snapshotting for Efficient Tactic Search in Lean 4

DGX agent

arXiv:2605.25556v1 Announce Type: cross Abstract: Automated theorem proving systems built on Lean 4 increasingly rely on parallel tactic search over partially specified proofs, such as those generated

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Kolmogorov-Arnold Fourier Networks

DGX agent

arXiv:2502.06018v3 Announce Type: replace-cross Abstract: Although Kolmogorov-Arnold-based interpretable networks (KANs) possess strong theoretical expressiveness, they suffer from severe parameter ex

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Latent Q-Barrier Shielding for Safe In-Context Reinforcement Learning

DGX agent

arXiv:2605.25267v1 Announce Type: cross Abstract: Safe in-context reinforcement learning (ICRL) adapts online from interaction history without test-time parameter updates while controlling episode cos

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Learning Fine-grained Parameter Sharing via Sparse Tensor Decomposition

DGX agent

arXiv:2411.09816v4 Announce Type: replace Abstract: Large neural networks achieve state-of-the-art performance on many tasks, yet their sheer size hinders deployment on resource-constrained devices. A

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Learning Sparse Compositional Functions with Norm-Constrained Neural Networks

DGX agent

arXiv:2605.25608v1 Announce Type: cross Abstract: The ability of deep neural networks to learn hierarchical features is widely regarded as a key mechanism underlying their success in high-dimensional

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Learning to Reason Efficiently with A* Post-Training

DGX agent

arXiv:2605.24597v1 Announce Type: new Abstract: Many applications of large language models (LLMs) require deductive reasoning, yet models frequently produce incorrect or redundant inference steps. We

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization

DGX agent

arXiv:2510.03827v2 Announce Type: replace-cross Abstract: LIBERO has emerged as a widely adopted benchmark for evaluating Vision-Language-Action (VLA) models; however, its current training and evaluat

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

LiveMCP-101: Stress Testing and Diagnosing MCP-enabled Agents on Challenging Queries

DGX agent

arXiv:2508.15760v2 Announce Type: replace-cross Abstract: Tool calling has emerged as a critical capability for AI agents. In contrast to conventional tool calling frameworks that rely on static, prov

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Llamion Technical Report

DGX agent

arXiv:2605.25676v1 Announce Type: new Abstract: We release Llamion, a family of 14B-parameter open-weight language models obtained by transforming Orion-14B into the standardized Llama-family architec

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

LLM-as-a-Reviewer: Benchmarking Their Ability, Divergence, and Prompt Injection Resistance as Paper Reviewers

DGX agent

arXiv:2605.25415v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in academic peer review, yet their reliability, alignment with human judgment, and robustness to adve

model-releasesarxiv-cs-cl
26 May 2026
← Previous
1…265266267268269…472
Next →