AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,603 results
26 May 2026

Efficient Long-Horizon Vision-Language-Action Models via Static-Dynamic Disentanglement

Model ReleasesDGX agent

arXiv:2602.03983v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for generalist robotic control. Built upon vision-language model (

EMA-Nesterov: Stabilizing Nesterov's Lookahead for Accelerated Deep Learning Optimization

Model ReleasesDGX agent

arXiv:2605.25395v1 Announce Type: new Abstract: Lookahead-based acceleration methods, such as Nesterov's momentum, are widely used in optimization, but they often become unreliable in deep learning tr

Emission-Aware Reinforcement Learning for Sustainable Electric Vehicle Charging and Carbon Dioxide Reduction Under Varying Renewable Penetration


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2605.24543v1 Announce Type: new Abstract: The rapid growth of Electric Vehicle (EV) adoption challenges power distribution networks through peak load spikes, voltage instability, and transformer

Emotional intelligence in large language models is fragmented across perception, cognition, and interaction

Model ReleasesDGX agent

arXiv:2605.24686v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly integrated into emotionally sensitive domains, the structural integrity of their emotional intelligence

Empirical Analysis and Detection of Hallucinations in LLM-Generated Bug Report Summaries

Model ReleasesDGX agent

arXiv:2605.24137v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to generate summaries of software bug reports, including sections such as Steps-to-Reproduce (S2R),

Enhancing Reliability in LLM-Based Secure Code Generation

Model ReleasesDGX agent

arXiv:2605.24300v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for code generation, but their security reliability remains inconsistent across languages and prompting s

Equation-Free Coarse Control of Distributed Parameter Systems via Local Neural Operators

Model ReleasesDGX agent

arXiv:2509.23975v2 Announce Type: replace-cross Abstract: The control of high-dimensional distributed parameter systems (DPS) remains a challenge when explicit coarse-grained equations are unavailable

ERNIE-Image Technical Report

Model ReleasesDGX agent

arXiv:2605.25347v1 Announce Type: cross Abstract: We introduce ERNIE-Image, an open-source text-to-image generation model built upon an 8B single-stream DiT architecture. ERNIE-Image aims to bridge th

everybody talks about the china->us catchup not enough people talking about the us-> china catchup great job @o_lacombe et al, @robert_mchar…

Model ReleasesDGX agent

everybody talks about the china->us catchup not enough people talking about the us-> china catchup great job @o_lacombe et al, @robert_mchardy et al! [AINews 3 Apr 2026] Gemma 4: The world's best smal

EvoCode-Bench: Evaluating Coding Agents in Multi-Turn Iterative Interactions

Model ReleasesDGX agent

arXiv:2605.24110v1 Announce Type: new Abstract: Coding agents are increasingly used as iterative development partners, but most benchmarks still evaluate one specification followed by one final assess

EvoEGF-Mol: Evolving Exponential Geodesic Flow for Structure-based Drug Design

Model ReleasesDGX agent

arXiv:2601.22466v2 Announce Type: replace Abstract: Structure-Based Drug Design (SBDD) aims to discover bioactive ligands. Conventional approaches construct probability paths separately in Euclidean a

Exploration of Perceptual Speech Features for Clinical Decision-Support in Mental Health Care

Model ReleasesDGX agent

arXiv:2605.24678v1 Announce Type: new Abstract: Speech and language technologies offer valuable opportunities for supporting mental health assessment through objective and interpretable cues. We prese

Explore Before You Solve: The Speed--Depth Trade-off in Epistemic Agents for ARC-AGI-3

Model ReleasesDGX agent

arXiv:2605.25931v1 Announce Type: new Abstract: We systematically investigate all 25 public ARC-AGI-3 games and find that every one is reachable through non-intelligent strategies: 10 in a single blin

Extending Embodied Question Answering from Perception to Decision

Model ReleasesDGX agent

arXiv:2605.25813v1 Announce Type: new Abstract: Embodied Question Answering (EQA) connects perception, reasoning, and interaction within embodied environments. However, existing datasets and benchmark

Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth

Model ReleasesDGX agent

arXiv:2605.25052v1 Announce Type: new Abstract: Chains of thought (CoTs) have become central in interpreting and auditing behaviors of large language models. Yet growing evidence suggests that these t

Feature Learning in Wide Neural Networks under muP: Identifiability and Sparse-Dictionary Decomposition of the Mean-Field Limit

Model ReleasesDGX agent

arXiv:2605.24710v1 Announce Type: new Abstract: We establish four structural results for feature learning in wide two-layer neural networks under the Maximal Update Parametrization (muP). First, we pr

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model

Model ReleasesDGX agent

arXiv:2510.10921v3 Announce Type: replace-cross Abstract: Fine-grained vision-language understanding requires precise alignment between visual content and linguistic descriptions, a capability that re

Fine-Tuning and Serving Gemma 4 31B on Google Cloud TPU: A Technical Comparison with GPU Baselines

Model ReleasesDGX agent

arXiv:2605.25645v1 Announce Type: cross Abstract: We present the first end-to-end demonstration of fine-tuning and serving Google's Gemma 4 31B model on TPU hardware, providing an empirical comparison

Fine-Tuning Language Models to Know What They Know

Model ReleasesDGX agent

arXiv:2602.02605v2 Announce Type: replace-cross Abstract: Evaluating true metacognition in Large Language Models (LLMs) is difficult due to biases and heuristics. This paper presents a framework to me

FLOATBench: A Dataset and Benchmark for Floating Offshore Wind Turbine Tower Fatigue

Model ReleasesDGX agent

arXiv:2605.25717v1 Announce Type: new Abstract: Most of the world's offshore wind resource lies in waters too deep for fixed-bottom foundations, making floating offshore wind turbines (FOWTs) essentia

FloorplanQA: A Benchmark for Spatial Reasoning in LLMs using Structured Representations

Model ReleasesDGX agent

arXiv:2507.07644v4 Announce Type: replace Abstract: We introduce FloorplanQA, a diagnostic benchmark for evaluating spatial reasoning in large language models (LLMs). FloorplanQA is grounded in struct

FLoRIST: Singular Value Thresholding for Efficient and Accurate Federated Fine-Tuning of Large Language Models

Model ReleasesDGX agent

arXiv:2506.09199v2 Announce Type: replace-cross Abstract: Integrating Low-Rank Adaptation (LoRA) into federated learning offers a promising solution for parameter-efficient fine-tuning of Large Langua

FoodMonitor: Benchmarking MLLMs for Explainable Compliance Analysis

Model ReleasesDGX agent

arXiv:2605.24503v1 Announce Type: cross Abstract: As AI-powered compliance monitoring becomes increasingly important in public governance and industrial safety, the ability to provide verifiable evide

Found in Conversation: LLMs Teach Themselves to Close the Multi-Turn Gap

Model ReleasesDGX agent

arXiv:2605.24432v1 Announce Type: new Abstract: Large Language Model (LLM) interactions are typically underspecified, with users clarifying all necessary details across multiple conversational turns.

FOUND-IT: Foundation-model-first Task-driven 3D Scene Graphs with Granularity on Demand

Model ReleasesDGX agent

arXiv:2605.25371v1 Announce Type: new Abstract: We present the first approach to build hierarchical task-driven 3D scene graphs of arbitrary indoor or outdoor environments using an uncalibrated monocu

Fourier Feature Pyramids for Physics-Informed Neural Networks

Model ReleasesDGX agent

arXiv:2605.24278v1 Announce Type: new Abstract: We present an improved neural field architecture for solving partial differential equations (PDEs). Current physics-informed neural networks (PINNs) pro

Free the 100B Gemma 4 MoE! Gemini Flash 3.5 is out so now you can release it!

Model ReleasesDGX agent

Clem Delangue advocates for the release of a 100 billion parameter Gemma 4 Mixture of Experts model, suggesting that Gemini Flash 3.5's release creates an opportunity for this larger model to be made

From DPPs to k-DPPs: identifiability analysis via spectral decomposition

Model ReleasesDGX agent

arXiv:2605.25526v1 Announce Type: cross Abstract: We study the geometry of determinantal point processes (DPPs) through the spectral decomposition L=ULambda U^{op}. The spectrum Lambda governs the car

From Facts to Insights: A Persona-Driven Dual Memory Framework and Dataset for Role-Playing Agents

Model ReleasesDGX agent

arXiv:2605.25693v1 Announce Type: new Abstract: While role-playing agents excel in short-term interactions, long-term conversations overwhelm context windows, motivating external memory frameworks. Cu

From Index to Equity: Pre-Training Transformers for Stock Return Prediction

Model ReleasesDGX agent

arXiv:2605.23962v1 Announce Type: cross Abstract: This research aims to leverage machine learning to improve stock price prediction and support informed investment decisions related to buying, selling

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

Model ReleasesDGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

From One-Pass SGD to Data Reuse: Mini-Batch Scaling Laws in Sketched Linear Regression

Model ReleasesDGX agent

arXiv:2605.24316v1 Announce Type: new Abstract: Scaling laws provide compact descriptions of how prediction error varies with compute, model size, and data, but existing theory mainly treats single-sa

From Prompt Optimization to Multi-Dimensional Credibility Evaluation: Enhancing Trustworthiness of Chinese LLM-Generated Liver MRI Reports -- with Preliminary Extension to Lung Cancer

Model ReleasesDGX agent

arXiv:2510.23008v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated promising performance in generating diagnostic conclusions from imaging findings, thereby supporting

From Theory to Decision Rule: Calibrating the Noisy-Label Crossover for Vision-Language Model Weak Supervision Across Three Medical-Imaging Benchmarks

Model ReleasesDGX agent

arXiv:2605.24771v1 Announce Type: cross Abstract: Classical noisy-label theory predicts that downstream performance under weak supervision is bounded above by the labeler's accuracy, implying a sharp

FrontierOR: Benchmarking LLMs' Capacity for Efficient Algorithm Design in Large-Scale Optimization

Model ReleasesDGX agent

arXiv:2605.25246v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for optimization modeling and solver-code generation, yet practical operations research and optimizat

Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model

Model ReleasesDGX agent

arXiv:2412.07333v2 Announce Type: replace-cross Abstract: Pose-Guided Person Image Synthesis (PGPIS) aims to generate human images in specified poses while preserving the identity and appearance of a

Game-Theoretic Modeling of Heterogeneous Investor Interactions for Stock Price Forecasting

Model ReleasesDGX agent

arXiv:2605.23953v1 Announce Type: cross Abstract: Accurate stock price forecasting has consistently remained a pivotal yet challenging FinTech task that underpins quantitative trading and investment d

GDformer: Going Beyond Subsequence Isolation for Multivariate Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2501.18196v3 Announce Type: replace Abstract: Unsupervised anomaly detection of multivariate time series is a challenging task, given the requirements of deriving a compact detection criterion w

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open …

Model ReleasesDGX agent

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open models. Some ideas for what comes next, May 2026 Gemini Flas

Generalizable Vision-Language Few-Shot Adaptation with Predictive Prompts and Negative Learning

Model ReleasesDGX agent

arXiv:2505.11758v2 Announce Type: replace-cross Abstract: Few-shot adaptation of vision-language models remains fundamentally limited by how negative class signals are handled at inference. Existing m

Geo-Expert: Towards Expert-Level Geological Reasoning via Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2605.24844v1 Announce Type: new Abstract: While general-purpose Large Language Models (LLMs) applied to Geology often hallucinate when reasoning about subsurface structures and deep-time evoluti

GL-LFGNN:A Global-Local Dual-branch Causal Graph Neural Network Based on Liang-Kleeman Information Flow for EEG Emotion Recognition

Model ReleasesDGX agent

arXiv:2605.25061v1 Announce Type: cross Abstract: EEG-based emotion recognition holds significant promise for objective diagnosis of mood disorders. Graph neural networks (GNNs) have emerged as the do

GlobalDentBench: A Multinational Benchmark for Evaluating LLM Clinical Reasoning in Dentistry with Expert Calibration

Model ReleasesDGX agent

arXiv:2605.24636v1 Announce Type: new Abstract: While large language models (LLMs) hold transformative potential for medicine, their reasoning robustness and safety in real-world clinical scenarios re

Goal-driven Bayesian Optimal Experimental Design for Robust Decision-Making Under Model Uncertainty

Model ReleasesDGX agent

arXiv:2605.26093v1 Announce Type: new Abstract: Bayesian optimal experimental design (BOED) selects experiments to maximize information gain about model parameters. However, in decision-critical setti

Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers

Model ReleasesDGX agent

arXiv:2605.24518v1 Announce Type: cross Abstract: The quadratic complexity of self-attention in Transformer models remains a significant bottleneck for processing long sequences and deploying large la

GreenSeg: Ground Segmentation Algorithm for Agricultural Robots in Mediterranean Greenhouses using RGB-D Point Clouds

Model ReleasesDGX agent

arXiv:2605.25279v1 Announce Type: new Abstract: Greenhouse agriculture in the Mediterranean region faces significant automation challenges due to its unique structural and environmental constraints. T

GroupTravelBench: Benchmarking LLM Agents on Multi-Person Travel Planning

Model ReleasesDGX agent

arXiv:2605.25200v1 Announce Type: new Abstract: Travel planning is a realistic task for evaluating the planning and tool-use abilities of LLM agents. However, existing benchmarks typically assume only

Guided Flow Matching for Forward and Inverse PDE Problems with Sparse Observations: Algorithm and Theory

Model ReleasesDGX agent

arXiv:2605.25509v1 Announce Type: cross Abstract: Reconstructing PDE solutions from sparse observations is a core challenge in scientific computing. We present FM4PDE, a flow-matching generative frame

Hadamard Representation: Scaffolding Performance Across Model-free RL

Model ReleasesDGX agent

arXiv:2406.09079v5 Announce Type: replace Abstract: Deep reinforcement learning agents progressively lose representational capacity during training: neurons become dormant, removing active capacity fr

HEAPr: Hessian-based Efficient Atomic Expert Pruning in Output Space

Model ReleasesDGX agent

arXiv:2509.22299v3 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) architectures in large language models (LLMs) deliver exceptional performance and reduced inference costs compared to

HiGraph: A Large-Scale Hierarchical Graph Dataset for Malware Analysis

Model ReleasesDGX agent

arXiv:2509.02113v2 Announce Type: replace-cross Abstract: The advancement of graph-based malware analysis is critically limited by the absence of large-scale datasets that capture the inherent hierarc

HiMed: Incentivizing Hindi Reasoning in Medical LLMs

Model ReleasesDGX agent

arXiv:2605.24635v1 Announce Type: new Abstract: Medical large language models hold promise for reducing healthcare disparities, yet Hindi remains severely underrepresented. While medical LLMs excel in

HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing

Model ReleasesDGX agent

arXiv:2605.24687v1 Announce Type: cross Abstract: Text-to-Image (T2I) models have made significant strides in visual realism and semantic consistency, yet they often perpetuate and amplify societal bi

How Many Tools Should an LLM Agent See? A Chance-Corrected Answer

Model ReleasesDGX agent

arXiv:2605.24660v1 Announce Type: cross Abstract: Before an LLM agent can use a tool, a retrieval system must decide which candidate tools to show to the agent. How long should that shortlist be? Show

How Much Do Large Language Model Cheat on Evaluation? Benchmarking Overestimation under the One-Time-Pad-Based Framework

Model ReleasesDGX agent

arXiv:2507.19219v2 Announce Type: replace Abstract: Overestimation in evaluating large language models (LLMs) has become an increasing concern. Due to the contamination of public benchmarks or imbalan

How Much Thinking is Enough? Quantifying and Understanding Redundancy in LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.23926v1 Announce Type: new Abstract: Reasoning-capable large language models solve hard problems by emitting long chains of thought, paying heavily in latency, GPU time, and energy. Casual

How we evolved Google’s global and data center networks for the AI era

Model ReleasesDGX agent

Over the last 25 years of building Google’s global network, we’ve navigated major architectural eras — from the Internet, to streaming, and the cloud. Today, we are squarely in the midst of a fourth:

How Well Do Models Follow Their Constitutions?

Model ReleasesDGX agent

arXiv:2605.24229v1 Announce Type: new Abstract: Frontier AI developers now train models against long written behavioral specifications, such as Anthropic's constitution (Anthropic, 2025a) and OpenAI's

Hylos: Operability Contracts for Model-Native Spatial Intelligence

Model ReleasesDGX agent

arXiv:2605.24728v1 Announce Type: new Abstract: Foundation models can increasingly describe, reconstruct, and generate 3D objects, assemblies, scenes, and environments, but visually plausible spatial

I don't comment on every article that uses out-of-date measures on AI ability, but I felt (probably wrongly) that the article was a response…

Model ReleasesDGX agent

I don't comment on every article that uses out-of-date measures on AI ability, but I felt (probably wrongly) that the article was a response to my viral tweet, so I felt I needed to say something! htt

← Previous
1…210211212213214…377
Next →