AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,797 results
Model Releases

EvoEGF-Mol: Evolving Exponential Geodesic Flow for Structure-based Drug Design

DGX agent

arXiv:2601.22466v2 Announce Type: replace Abstract: Structure-Based Drug Design (SBDD) aims to discover bioactive ligands. Conventional approaches construct probability paths separately in Euclidean a

model-releasesarxiv-cs-lg
26 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Exploration of Perceptual Speech Features for Clinical Decision-Support in Mental Health Care

DGX agent

arXiv:2605.24678v1 Announce Type: new Abstract: Speech and language technologies offer valuable opportunities for supporting mental health assessment through objective and interpretable cues. We prese

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Explore Before You Solve: The Speed--Depth Trade-off in Epistemic Agents for ARC-AGI-3

DGX agent

arXiv:2605.25931v1 Announce Type: new Abstract: We systematically investigate all 25 public ARC-AGI-3 games and find that every one is reachable through non-intelligent strategies: 10 in a single blin

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Extending Embodied Question Answering from Perception to Decision

DGX agent

arXiv:2605.25813v1 Announce Type: new Abstract: Embodied Question Answering (EQA) connects perception, reasoning, and interaction within embodied environments. However, existing datasets and benchmark

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth

DGX agent

arXiv:2605.25052v1 Announce Type: new Abstract: Chains of thought (CoTs) have become central in interpreting and auditing behaviors of large language models. Yet growing evidence suggests that these t

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Feature Learning in Wide Neural Networks under muP: Identifiability and Sparse-Dictionary Decomposition of the Mean-Field Limit

DGX agent

arXiv:2605.24710v1 Announce Type: new Abstract: We establish four structural results for feature learning in wide two-layer neural networks under the Maximal Update Parametrization (muP). First, we pr

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model

DGX agent

arXiv:2510.10921v3 Announce Type: replace-cross Abstract: Fine-grained vision-language understanding requires precise alignment between visual content and linguistic descriptions, a capability that re

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Fine-Tuning and Serving Gemma 4 31B on Google Cloud TPU: A Technical Comparison with GPU Baselines

DGX agent

arXiv:2605.25645v1 Announce Type: cross Abstract: We present the first end-to-end demonstration of fine-tuning and serving Google's Gemma 4 31B model on TPU hardware, providing an empirical comparison

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Fine-Tuning Language Models to Know What They Know

DGX agent

arXiv:2602.02605v2 Announce Type: replace-cross Abstract: Evaluating true metacognition in Large Language Models (LLMs) is difficult due to biases and heuristics. This paper presents a framework to me

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

FLOATBench: A Dataset and Benchmark for Floating Offshore Wind Turbine Tower Fatigue

DGX agent

arXiv:2605.25717v1 Announce Type: new Abstract: Most of the world's offshore wind resource lies in waters too deep for fixed-bottom foundations, making floating offshore wind turbines (FOWTs) essentia

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

FloorplanQA: A Benchmark for Spatial Reasoning in LLMs using Structured Representations

DGX agent

arXiv:2507.07644v4 Announce Type: replace Abstract: We introduce FloorplanQA, a diagnostic benchmark for evaluating spatial reasoning in large language models (LLMs). FloorplanQA is grounded in struct

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

FLoRIST: Singular Value Thresholding for Efficient and Accurate Federated Fine-Tuning of Large Language Models

DGX agent

arXiv:2506.09199v2 Announce Type: replace-cross Abstract: Integrating Low-Rank Adaptation (LoRA) into federated learning offers a promising solution for parameter-efficient fine-tuning of Large Langua

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

FoodMonitor: Benchmarking MLLMs for Explainable Compliance Analysis

DGX agent

arXiv:2605.24503v1 Announce Type: cross Abstract: As AI-powered compliance monitoring becomes increasingly important in public governance and industrial safety, the ability to provide verifiable evide

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Found in Conversation: LLMs Teach Themselves to Close the Multi-Turn Gap

DGX agent

arXiv:2605.24432v1 Announce Type: new Abstract: Large Language Model (LLM) interactions are typically underspecified, with users clarifying all necessary details across multiple conversational turns.

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

FOUND-IT: Foundation-model-first Task-driven 3D Scene Graphs with Granularity on Demand

DGX agent

arXiv:2605.25371v1 Announce Type: new Abstract: We present the first approach to build hierarchical task-driven 3D scene graphs of arbitrary indoor or outdoor environments using an uncalibrated monocu

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Fourier Feature Pyramids for Physics-Informed Neural Networks

DGX agent

arXiv:2605.24278v1 Announce Type: new Abstract: We present an improved neural field architecture for solving partial differential equations (PDEs). Current physics-informed neural networks (PINNs) pro

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Free the 100B Gemma 4 MoE! Gemini Flash 3.5 is out so now you can release it!

DGX agent

Clem Delangue advocates for the release of a 100 billion parameter Gemma 4 Mixture of Experts model, suggesting that Gemini Flash 3.5's release creates an opportunity for this larger model to be made

model-releasesclem-delangue--x
26 May 2026
Model Releases

From DPPs to k-DPPs: identifiability analysis via spectral decomposition

DGX agent

arXiv:2605.25526v1 Announce Type: cross Abstract: We study the geometry of determinantal point processes (DPPs) through the spectral decomposition L=ULambda U^{op}. The spectrum Lambda governs the car

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

From Facts to Insights: A Persona-Driven Dual Memory Framework and Dataset for Role-Playing Agents

DGX agent

arXiv:2605.25693v1 Announce Type: new Abstract: While role-playing agents excel in short-term interactions, long-term conversations overwhelm context windows, motivating external memory frameworks. Cu

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

From Index to Equity: Pre-Training Transformers for Stock Return Prediction

DGX agent

arXiv:2605.23962v1 Announce Type: cross Abstract: This research aims to leverage machine learning to improve stock price prediction and support informed investment decisions related to buying, selling

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

DGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

From One-Pass SGD to Data Reuse: Mini-Batch Scaling Laws in Sketched Linear Regression

DGX agent

arXiv:2605.24316v1 Announce Type: new Abstract: Scaling laws provide compact descriptions of how prediction error varies with compute, model size, and data, but existing theory mainly treats single-sa

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

From Prompt Optimization to Multi-Dimensional Credibility Evaluation: Enhancing Trustworthiness of Chinese LLM-Generated Liver MRI Reports -- with Preliminary Extension to Lung Cancer

DGX agent

arXiv:2510.23008v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated promising performance in generating diagnostic conclusions from imaging findings, thereby supporting

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

From Theory to Decision Rule: Calibrating the Noisy-Label Crossover for Vision-Language Model Weak Supervision Across Three Medical-Imaging Benchmarks

DGX agent

arXiv:2605.24771v1 Announce Type: cross Abstract: Classical noisy-label theory predicts that downstream performance under weak supervision is bounded above by the labeler's accuracy, implying a sharp

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

FrontierOR: Benchmarking LLMs' Capacity for Efficient Algorithm Design in Large-Scale Optimization

DGX agent

arXiv:2605.25246v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for optimization modeling and solver-code generation, yet practical operations research and optimizat

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model

DGX agent

arXiv:2412.07333v2 Announce Type: replace-cross Abstract: Pose-Guided Person Image Synthesis (PGPIS) aims to generate human images in specified poses while preserving the identity and appearance of a

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Game-Theoretic Modeling of Heterogeneous Investor Interactions for Stock Price Forecasting

DGX agent

arXiv:2605.23953v1 Announce Type: cross Abstract: Accurate stock price forecasting has consistently remained a pivotal yet challenging FinTech task that underpins quantitative trading and investment d

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

GDformer: Going Beyond Subsequence Isolation for Multivariate Time Series Anomaly Detection

DGX agent

arXiv:2501.18196v3 Announce Type: replace Abstract: Unsupervised anomaly detection of multivariate time series is a challenging task, given the requirements of deriving a compact detection criterion w

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open …

DGX agent

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open models. Some ideas for what comes next, May 2026 Gemini Flas

model-releasesclem-delangue--x
26 May 2026
Model Releases

Generalizable Vision-Language Few-Shot Adaptation with Predictive Prompts and Negative Learning

DGX agent

arXiv:2505.11758v2 Announce Type: replace-cross Abstract: Few-shot adaptation of vision-language models remains fundamentally limited by how negative class signals are handled at inference. Existing m

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Geo-Expert: Towards Expert-Level Geological Reasoning via Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2605.24844v1 Announce Type: new Abstract: While general-purpose Large Language Models (LLMs) applied to Geology often hallucinate when reasoning about subsurface structures and deep-time evoluti

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

GL-LFGNN:A Global-Local Dual-branch Causal Graph Neural Network Based on Liang-Kleeman Information Flow for EEG Emotion Recognition

DGX agent

arXiv:2605.25061v1 Announce Type: cross Abstract: EEG-based emotion recognition holds significant promise for objective diagnosis of mood disorders. Graph neural networks (GNNs) have emerged as the do

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

GlobalDentBench: A Multinational Benchmark for Evaluating LLM Clinical Reasoning in Dentistry with Expert Calibration

DGX agent

arXiv:2605.24636v1 Announce Type: new Abstract: While large language models (LLMs) hold transformative potential for medicine, their reasoning robustness and safety in real-world clinical scenarios re

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Goal-driven Bayesian Optimal Experimental Design for Robust Decision-Making Under Model Uncertainty

DGX agent

arXiv:2605.26093v1 Announce Type: new Abstract: Bayesian optimal experimental design (BOED) selects experiments to maximize information gain about model parameters. However, in decision-critical setti

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers

DGX agent

arXiv:2605.24518v1 Announce Type: cross Abstract: The quadratic complexity of self-attention in Transformer models remains a significant bottleneck for processing long sequences and deploying large la

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

GreenSeg: Ground Segmentation Algorithm for Agricultural Robots in Mediterranean Greenhouses using RGB-D Point Clouds

DGX agent

arXiv:2605.25279v1 Announce Type: new Abstract: Greenhouse agriculture in the Mediterranean region faces significant automation challenges due to its unique structural and environmental constraints. T

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

GroupTravelBench: Benchmarking LLM Agents on Multi-Person Travel Planning

DGX agent

arXiv:2605.25200v1 Announce Type: new Abstract: Travel planning is a realistic task for evaluating the planning and tool-use abilities of LLM agents. However, existing benchmarks typically assume only

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Guided Flow Matching for Forward and Inverse PDE Problems with Sparse Observations: Algorithm and Theory

DGX agent

arXiv:2605.25509v1 Announce Type: cross Abstract: Reconstructing PDE solutions from sparse observations is a core challenge in scientific computing. We present FM4PDE, a flow-matching generative frame

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Hadamard Representation: Scaffolding Performance Across Model-free RL

DGX agent

arXiv:2406.09079v5 Announce Type: replace Abstract: Deep reinforcement learning agents progressively lose representational capacity during training: neurons become dormant, removing active capacity fr

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

HEAPr: Hessian-based Efficient Atomic Expert Pruning in Output Space

DGX agent

arXiv:2509.22299v3 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) architectures in large language models (LLMs) deliver exceptional performance and reduced inference costs compared to

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

HiGraph: A Large-Scale Hierarchical Graph Dataset for Malware Analysis

DGX agent

arXiv:2509.02113v2 Announce Type: replace-cross Abstract: The advancement of graph-based malware analysis is critically limited by the absence of large-scale datasets that capture the inherent hierarc

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

HiMed: Incentivizing Hindi Reasoning in Medical LLMs

DGX agent

arXiv:2605.24635v1 Announce Type: new Abstract: Medical large language models hold promise for reducing healthcare disparities, yet Hindi remains severely underrepresented. While medical LLMs excel in

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing

DGX agent

arXiv:2605.24687v1 Announce Type: cross Abstract: Text-to-Image (T2I) models have made significant strides in visual realism and semantic consistency, yet they often perpetuate and amplify societal bi

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

How Many Tools Should an LLM Agent See? A Chance-Corrected Answer

DGX agent

arXiv:2605.24660v1 Announce Type: cross Abstract: Before an LLM agent can use a tool, a retrieval system must decide which candidate tools to show to the agent. How long should that shortlist be? Show

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

How Much Do Large Language Model Cheat on Evaluation? Benchmarking Overestimation under the One-Time-Pad-Based Framework

DGX agent

arXiv:2507.19219v2 Announce Type: replace Abstract: Overestimation in evaluating large language models (LLMs) has become an increasing concern. Due to the contamination of public benchmarks or imbalan

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

How Much Thinking is Enough? Quantifying and Understanding Redundancy in LLM Reasoning

DGX agent

arXiv:2605.23926v1 Announce Type: new Abstract: Reasoning-capable large language models solve hard problems by emitting long chains of thought, paying heavily in latency, GPU time, and energy. Casual

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

How we evolved Google’s global and data center networks for the AI era

DGX agent

Over the last 25 years of building Google’s global network, we’ve navigated major architectural eras — from the Internet, to streaming, and the cloud. Today, we are squarely in the midst of a fourth:

model-releasesgoogle-cloud-ai
26 May 2026
Model Releases

How Well Do Models Follow Their Constitutions?

DGX agent

arXiv:2605.24229v1 Announce Type: new Abstract: Frontier AI developers now train models against long written behavioral specifications, such as Anthropic's constitution (Anthropic, 2025a) and OpenAI's

model-releasesarxiv-cs-ai
26 May 2026
← Previous
1…267268269270271…475
Next →