AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,603 results
25 May 2026

Joint Model Parameter Scaling and Universal-Domain Data Integration for E-commerce Search Ranking

Model ReleasesDGX agent

arXiv:2603.24226v3 Announce Type: replace-cross Abstract: Scaling studies for industrial search, advertising, and recommendation have largely emphasized enlarging model capacity or refining architectu

Just found that if you scroll down in the Claude Code app on iPhone… Clawd starts jumping and walking around looking for apps to build

Model ReleasesDGX agent

The Claude Code app for iPhone features an interactive Easter egg where scrolling down triggers an animated character named 'Clawd' that jumps and walks around the screen, appearing to search for apps

Just spent a full coding session with Grok Build and honestly? It's right there with Claude. The model is sharp, the agentic flow holds up o…


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

Just spent a full coding session with Grok Build and honestly? It's right there with Claude. The model is sharp, the agentic flow holds up on complex tasks, and it has actual personality. Few rough ed

KAPLAN: Kolmogorov-Arnold Prognostic Learnable Activation Networks for Survival Analysis

Model ReleasesDGX agent

arXiv:2605.23082v1 Announce Type: cross Abstract: Survival analysis aims to model how covariates and time jointly shape the time-to-event distribution under right censoring. Classical methods such as

L-FAME: Longitudinal Focused Attention Meditation EEG Dataset and Benchmark

Model ReleasesDGX agent

arXiv:2605.22893v1 Announce Type: cross Abstract: We introduce a novel Longitudinal Focused Attention Meditation Electroencephalography (L-FAME) dataset and an accompanying benchmark, designed to fost

Learning Safely Without Knowing the World:COMPASS-Hedge

Model ReleasesDGX agent

arXiv:2603.22348v3 Announce Type: replace Abstract: Online learning algorithms often face a fundamental trilemma: balancing regret guarantees between adversarial and stochastic settings and providing

LFRAG: Layout-oriented Fine-grained Retrieval-Augmented Generation on Multimodal Document Understanding

Model ReleasesDGX agent

arXiv:2605.22829v1 Announce Type: cross Abstract: Multimodal Retrieval-Augmented Generation (RAG) has emerged as an effective paradigm for enhancing Large Language Models (LLMs) with external knowledg

Lipschitz Optimization for Formal Verification of Homographies

Model ReleasesDGX agent

arXiv:2605.23203v1 Announce Type: cross Abstract: The adoption of vision neural networks in regulated industries requires formal robustness guarantees, especially in safety-critical domains such as he

LLAMA LIMA: A Living Meta-Analysis on the Effects of Generative AI on Learning Mathematics

Model ReleasesDGX agent

arXiv:2601.18685v3 Announce Type: replace-cross Abstract: The capabilities of generative AI in mathematics education are rapidly evolving, posing significant challenges for research to keep pace. Rese

LLM-driven design of physics-constrained constitutive models: two agents are better than one

Model ReleasesDGX agent

arXiv:2605.23754v1 Announce Type: new Abstract: Developing constitutive models that capture how materials deform under load traditionally requires years of specialized expertise in continuum mechanics

LQ-rPPG: A Label-Quantized Coarse-to-Fine Learning Framework for Remote Physiological Measurement

Model ReleasesDGX agent

arXiv:2605.23174v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) enables non-contact measurement of physiological signals from facial videos, offering strong potential for remote hea

MadEvolve: Evolutionary Optimization of Trading Systems with Large Language Models

Model ReleasesDGX agent

arXiv:2605.23007v1 Announce Type: cross Abstract: We explore the application of LLM-driven algorithm optimization to several common tasks in quantitative finance. MadEvolve, a general-purpose algorith

MAS-Orchestra: Understanding and Improving Multi-Agent Reasoning Through Holistic Orchestration and Controlled Benchmarks

Model ReleasesDGX agent

arXiv:2601.14652v5 Announce Type: replace Abstract: While multi-agent systems (MAS) promise elevated intelligence through coordination of agents, current approaches to automatic MAS design under-deliv

MedExpMem: Adapting Experience Memory for Differential Diagnosis

Model ReleasesDGX agent

arXiv:2605.22872v1 Announce Type: cross Abstract: Experienced physicians develop diagnostic expertise through clinical practice, acquiring not only disease knowledge but also the ability to differenti

MELT: A Behavioral Trace Dataset for High-Risk Memecoin Launch Detection

Model ReleasesDGX agent

arXiv:2602.13480v2 Announce Type: cross Abstract: Launchpads have become the dominant mechanism for issuing memecoins, exposing investors to a new class of high-risk launches that existing rug-pull de

Memorization Dynamics of Fill-in-the-Middle Pretraining

Model ReleasesDGX agent

arXiv:2605.22981v1 Announce Type: cross Abstract: Fill-in-the-middle (FIM) is a pretraining objective widely used to equip causal language models with infilling ability, yet its effect on verbatim mem

Metadata Predictability Is Not Evidence Dependence: An Intervention-Based Audit for Weak-Label Benchmarks

Model ReleasesDGX agent

arXiv:2605.23701v1 Announce Type: new Abstract: We study a protocol-level test for weak-label benchmarks: whether benchmark outputs change when the provided evidence is intervened on. Metadata-only sh

Model Collapse as Cultural Evolution

Model ReleasesDGX agent

arXiv:2605.23054v1 Announce Type: cross Abstract: Model collapse, the progressive degradation of LLMs trained on their own outputs, has been characterized statistically but lacks a linguistic explanat

ModeSwitch-LLM: A Lightweight Phase-Aware Controller for Cross-Mode LLM Inference on a Single GPU

Model ReleasesDGX agent

arXiv:2605.23057v1 Announce Type: cross Abstract: ModeSwitch-LLM is a lightweight request-boundary controller for improving single-GPU large language model inference efficiency by routing each request

Moonwalk: Inverse-Forward Differentiation

Model ReleasesDGX agent

arXiv:2402.14212v4 Announce Type: replace-cross Abstract: Backpropagation's main limitation is its need to store intermediate activations (residuals) during the forward pass, which restricts the depth

Move on Muon : A Hamiltonian probability gradient flow perspective of Muon optimizer

Model ReleasesDGX agent

arXiv:2605.23871v1 Announce Type: cross Abstract: We develop a gradient flow on the space of probability measures defined on matrix-valued parameters induced by regularized Muon, an analytically smoot

Multi-SpatialMLLM: Multi-Frame Spatial Understanding with Multi-Modal Large Language Models

Model ReleasesDGX agent

arXiv:2505.17015v2 Announce Type: replace-cross Abstract: Multi-modal large language models (MLLMs) have rapidly advanced in visual tasks, yet their spatial understanding remains limited to single ima

Multilingual Steering by Design: Multilingual Sparse Autoencoders and Principled Layer Selection

Model ReleasesDGX agent

arXiv:2605.23036v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) enable feature-level mechanistic interpretability and activation steering in large language models (LLMs), but SAE-based lang

Non-normal spectral signatures of instability in neural network training dynamics

Model ReleasesDGX agent

arXiv:2605.23476v1 Announce Type: new Abstract: Training instabilities in deep networks - loss spikes, oscillatory convergence, and gradient pathologies - are empirically prevalent but lack a rigorous

NP-LoRA: Null Space Projection for Subject-Style LoRA Fusion

Model ReleasesDGX agent

arXiv:2511.11051v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) fusion enables the composition of subject and style representations for controllable generation without retraining. Howev

One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents

Model ReleasesDGX agent

arXiv:2605.23652v1 Announce Type: new Abstract: On a 300-persona life-simulation benchmark, pcsp achieves compositional zero-shot persona identification up to 17x above chance, Spearman rho approx 0.7

Online Partitioned Local Depth for semi-supervised applications

Model ReleasesDGX agent

arXiv:2512.15436v2 Announce Type: replace-cross Abstract: We introduce an extension of the partitioned local depth (PaLD) algorithm that is adapted to online applications such as semi-supervised predi

Ontological Knowledge Blocks: Executable Compliance and Profile-Based Validation for Trustworthy AI Systems

Model ReleasesDGX agent

arXiv:2605.23297v1 Announce Type: new Abstract: AI-enabled services deployed in critical digital infrastructure are subject to governance obligations spanning transparency, accountability, fairness, a

Open Multimodal Datasets and Open-Source Software for Data-Driven Modeling of Multiphase Transport and Thermal Systems

Model ReleasesDGX agent

arXiv:2605.23037v1 Announce Type: new Abstract: Data-driven modeling is becoming central to multiphase transport, electronics cooling, acoustic diagnostics, and thermal-fluid digital twins, but progre

OpenAI, Grupo Folha and Grupo UOL announce strategic content partnership

Model ReleasesDGX agent

OpenAI announced a strategic partnership with Grupo Folha and Grupo UOL, major Brazilian media companies, to integrate their content into OpenAI's AI models and products. The partnership enables OpenA

OpenSkillEval: Automatically Auditing the Open Skill Ecosystem for LLM Agents

Model ReleasesDGX agent

arXiv:2605.23657v1 Announce Type: new Abstract: Skills, i.e., structured workflow instructions distilled for large language models (LLMs), are becoming an increasingly important mechanism for improvin

Operator Learning for Reconstructing Flow Fields from Sparse Measurements: a Language Model Approach

Model ReleasesDGX agent

arXiv:2605.23712v1 Announce Type: cross Abstract: Reconstructing flow fields from sparse measurements is a fundamental problem in fluid mechanics with broad implications for modeling, control, and des

Optimization of randomized neural networks for transfer operator approximation

Model ReleasesDGX agent

arXiv:2605.23689v1 Announce Type: new Abstract: RaNNDy is a randomized neural network architecture for the data-driven approximation of transfer operators associated with complex dynamical systems. Th

Order-Optimal Sequential 1-Bit Mean Estimation in General Tail Regimes

Model ReleasesDGX agent

arXiv:2604.07796v2 Announce Type: replace-cross Abstract: In this paper, we study the problem of mean estimation under 1-bit communication constraints. We propose a novel adaptive mean estimator based

PACE: Two-Timescale Self-Evolution for Small Language Model Agents

Model ReleasesDGX agent

arXiv:2605.23019v1 Announce Type: new Abstract: Deploying language-model agents in production often requires substantial compute and human effort to tune prompts, parsers, validators, and other compon

Parallel Context Compaction for Long-Horizon LLM Agent Serving

Model ReleasesDGX agent

arXiv:2605.23296v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate growing conversation histories that eventually exceed the model's context window. Context compaction via LLM-based su

PGT: Procedurally Generated Tasks for improving visual grounding in MLLMs

Model ReleasesDGX agent

arXiv:2605.23883v1 Announce Type: cross Abstract: Despite remarkable progress in Multimodal Large Language Models (MLLMs), these models still struggle with fine-grained understanding tasks. In this wo

Philosophical Dispositions as Behavioral Constraints for AI-Assisted Code Review: An Empirical Study

Model ReleasesDGX agent

arXiv:2605.23108v1 Announce Type: cross Abstract: AI-assisted code review tools typically operate as generic 'expert reviewer' agents, producing homogeneous findings regardless of the analysis type ne

PhotoFlow: Agentic 3D Virtual Photography Missions

Model ReleasesDGX agent

arXiv:2605.23771v1 Announce Type: cross Abstract: Virtual photography asks an agent to enter a prepared 3D scene with no preselected camera pose or reference image, infer a suitable shot from scene in

Physics-Informed Machine Learning Regulated by Finite Element Analysis for Simulation Acceleration of Melt Pool Dynamics in Laser Powder Bed Fusion

Model ReleasesDGX agent

arXiv:2506.20537v3 Announce Type: replace Abstract: Efficient simulation of Laser Powder Bed Fusion (LPBF) is crucial for process prediction due to the lasting issue of high computational cost associa

Physiome-ODE: A Benchmark for Irregularly Sampled Multivariate Time Series Forecasting Based on Biological ODEs

Model ReleasesDGX agent

arXiv:2502.07489v2 Announce Type: replace Abstract: State-of-the-art methods for forecasting irregularly sampled time series with missing values predominantly rely on just four datasets and a few smal

PixelPonder: Dynamic Patch Adaptation for Enhanced Multi-Conditional Text-to-Image Generation

Model ReleasesDGX agent

arXiv:2503.06684v3 Announce Type: replace Abstract: Recent advances in diffusion-based text-to-image generation have demonstrated promising results through visual condition control. However, existing

Pointwise Metrics Mislead: An Evaluation Protocol for Multimodal Inverse Problems

Model ReleasesDGX agent

arXiv:2605.22891v1 Announce Type: new Abstract: Evaluation in scientific reconstruction is dominated by pointwise metrics - RMSE, MAE, per-event resolution - under the implicit assumption that lower e

PoisonForge: Task-Level Targeted Poisoning Benchmark for Instruction-Tuned LLMs

Model ReleasesDGX agent

arXiv:2605.23168v1 Announce Type: cross Abstract: When practitioners fine-tune LLMs on unvetted datasets, an adversary can exploit the data supply chain through task-level poisoning: inserting a small

Pope Leo calls for being ‘profoundly human’ in the age of AI

Model ReleasesDGX agent

Pope Leo XIV warned of the risks of AI and unconstrained technological power in his first major papal document released on Monday. Magnifica Humanitas is the pope's manifesto on 'safeguarding the huma

Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks

Model ReleasesDGX agent

arXiv:2605.23170v1 Announce Type: cross Abstract: Position-controlled evaluation is standard for retrieval tasks such as Needle-in-a-Haystack and RULER, but mainstream reasoning benchmarks do not cont

PrefBench: Evaluating Zero-Shot LLM Agents in Hidden-Preference Personalized Pricing Negotiations

Model ReleasesDGX agent

arXiv:2605.22855v1 Announce Type: cross Abstract: Personalized pricing negotiations are a challenging testbed for LLM agents because successful interaction does not guarantee profitable decision makin

PROGRESSLM: Towards Progress Reasoning in Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.15224v2 Announce Type: replace-cross Abstract: Estimating task progress requires reasoning over long-horizon dynamics rather than recognizing static visual content. While modern Vision-Lang

ProtDBench: A Unified Benchmark of Protein Binder Design and Evaluation

Model ReleasesDGX agent

arXiv:2605.04118v2 Announce Type: replace-cross Abstract: Recent advances in de novo protein binder design have enabled increasing experimental validation, yet reported in silico metrics remain diffic

Push Your Agent: Measuring and Enforcing Quantitative Goal Persistence in Long-Horizon LLM Agents

Model ReleasesDGX agent

arXiv:2605.23574v1 Announce Type: new Abstract: Long-horizon language agents can make many plausible local tool calls yet fail to persist until a requested count is actually complete. We study this ga

R^3L: Reflect-then-Retry Reinforcement Learning with Language-Guided Exploration, Pivotal Credit, and Positive Amplification

Model ReleasesDGX agent

arXiv:2601.03715v2 Announce Type: replace-cross Abstract: Reinforcement learning drives recent advances in LLM reasoning and agentic capabilities, yet current approaches struggle with both exploration

Recursive Block-Diagonal Coupling for Resource-Efficient Training of Vision Models

Model ReleasesDGX agent

arXiv:2605.23656v1 Announce Type: new Abstract: Training high-capacity vision models from scratch requires substantial computational resources. To improve training efficiency of a wide target model, e

Reinforcement Learning for Microcanonical Graph Ensemble with Assortativity Constraints

Model ReleasesDGX agent

arXiv:2605.23285v1 Announce Type: cross Abstract: How network structure determines function is a fundamental question, and it can be investigated by graph ensembles with precisely controlled structura

Resilience Characterization of AI-Native Wireless Receivers via Persistent Homology

Model ReleasesDGX agent

arXiv:2605.22886v1 Announce Type: cross Abstract: AI-native wireless receivers based on deep learning exhibit remarkable performance under stationary channel conditions, yet their resilience to distri

Revitalizing Dense Material Segmentation: Stabilized Vision Transformers and the Generalization Paradox

Model ReleasesDGX agent

arXiv:2605.23747v1 Announce Type: new Abstract: Material segmentation, the pixel-wise classification of physical surface properties, remains a challenging problem in computer vision, requiring physico

RMA: an Agentic System for Research-Level Mathematical Problems

Model ReleasesDGX agent

arXiv:2605.22875v1 Announce Type: new Abstract: We present extbf{Research Math Agents (RMA)}, an agentic framework for automated reasoning on research-level mathematical problems. Unlike prior studies

RoboSurg-VQA: A Multimodal Benchmark for Surgical Segmentation-Aware Visual Question Answering

Model ReleasesDGX agent

arXiv:2605.23068v1 Announce Type: new Abstract: Reliable visual understanding in robot-assisted and minimally invasive surgery (RMIS/MIS) demands more than accurate masks: in clinical practice, clinic

Same Model, Different Weakness: How Language and Modality Reshape the Jailbreak Attack Surface in Frontier MLLMs

Model ReleasesDGX agent

arXiv:2605.23157v1 Announce Type: new Abstract: The attack surface of a multimodal large language model (MLLM) is language-dependent in ways that reveal the mechanistic structure of alignment failures

SciAtlas: A Large-Scale Knowledge Graph for Automated Scientific Research

Model ReleasesDGX agent

arXiv:2605.22878v1 Announce Type: new Abstract: The exponential growth of global academic output has confronted researchers and AI agents with an unprecedented ``information explosion,'' where fragmen

SciHorizon-GENE: Benchmarking LLM for Life Sciences Inference from Gene Knowledge to Functional Understanding

Model ReleasesDGX agent

arXiv:2601.12805v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown growing promise in biomedical research, particularly for knowledge-driven interpretation tasks. Howeve

← Previous
1…217218219220221…377
Next →