AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,624 results
12 May 2026

Composing diffusion priors with explicit physical context via generative Gibbs sampling

ResearchDGX agent

arXiv:2605.10642v1 Announce Type: new Abstract: Pretrained diffusion models provide powerful learned priors, but in scientific sampling the target distribution often depends on physical context that i

Computer Use at the Edge of the Statistical Precipice

Model ReleasesDGX agent

arXiv:2605.08261v1 Announce Type: cross Abstract: Evaluating Computer Use Agents (CUAs) on interactive environments is fraught with methodological pitfalls that the field has yet to systematically add

Convergence Analysis of Newton's Method for Neural Networks in the Overparameterized Limit

Model ReleasesDGX agent

arXiv:2605.08352v1 Announce Type: new Abstract: A convergence analysis is developed for the regularized Newton method for training neural networks (NNs) in the overparameterized limit. As the number o

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cross-Sample Relational Fusion: Unifying Domain Generalization and Class-Incremental Learning

Model ReleasesDGX agent

arXiv:2605.08839v1 Announce Type: new Abstract: Class-Incremental Learning (CIL) requires a learning system to learn new classes while retaining previously learned knowledge. However, in real-world sc

CUDABeaver: Benchmarking LLM-Based Automated CUDA Debugging

Model ReleasesDGX agent

arXiv:2605.08455v1 Announce Type: new Abstract: Debugging CUDA programs has long been challenging because failures often arise from subtle interactions among hardware behavior, compiler decisions, mem

DGPO: Beyond Pairwise Preferences with Directional Consistent Groupwise Optimization

SafetyDGX agent

arXiv:2605.10863v1 Announce Type: new Abstract: Although Large Language Models (LLMs) have made remarkable progress, current preference optimization methods still struggle to align directional consist

Discovery of Nonlinear Dynamics with Automated Basis Function Generation

ResearchDGX agent

arXiv:2605.09696v1 Announce Type: new Abstract: Discovering governing equations from observational data remains a fundamental challenge in scientific modeling, particularly when the underlying mathema

Discrete Flow Matching: Convergence Guarantees Under Minimal Assumptions

ResearchDGX agent

arXiv:2605.08882v1 Announce Type: new Abstract: Flow Matching has recently emerged as a popular class of generative models for simulating a target distribution mu_1 from samples drawn from a source di

Do LLMs Experience an Internal Polylogue? Investigating Reasoning through the Lens of Personas

TutorialsDGX agent

arXiv:2605.09159v1 Announce Type: new Abstract: Recent work shows that large language models (LLMs) encode behavioural traits ('personas') as linear directions in activation space, often called 'perso

Do not copy and paste! Rewriting strategies for code retrieval

Model ReleasesDGX agent

arXiv:2605.08299v1 Announce Type: cross Abstract: Embedding-based code retrieval often suffers when encoders overfit to surface syntax. Prior work mitigates this by using LLMs to rephrase queries and

Dynamics-Aligned Shared Hypernetworks for Contextual RL under Discontinuous Shifts

Model ReleasesDGX agent

arXiv:2602.06550v2 Announce Type: replace-cross Abstract: Zero-shot generalization in contextual reinforcement learning remains a core challenge, particularly when the context is latent and must be in

EconWebArena: Benchmarking Autonomous Agents on Economic Tasks in Realistic Web Environments

Model ReleasesDGX agent

arXiv:2506.08136v3 Announce Type: replace Abstract: We introduce EconWebArena, a benchmark for evaluating autonomous agents on complex, multimodal economic tasks in realistic web environments. The ben

Entropy-informed Decoding: Adaptive Information-Driven Branching

ResearchDGX agent

arXiv:2605.09745v1 Announce Type: cross Abstract: Large language models (LLMs) achieve remarkable generative performance, yet their output quality is dependent on the decoding strategy. While sampling

Evaluating Developmental Cognition Capabilities of LLMs

ResearchDGX agent

arXiv:2605.08549v1 Announce Type: new Abstract: Conversational AI is increasingly personalized around users' preferences, histories, goals, and knowledge, but much less around how users interpret and

Even @haider1 sees that Mythos has been overhyped.

Model ReleasesDGX agent

Even @haider1 sees that Mythos has been overhyped. mythos is pretty on par with gpt-5.5 and while gpt-5.5 is currently SOTA, it's not anything like what anthropic describes mythos as it's pretty obvio

Exactness Matters for Physical Rule Enforcement

Model ReleasesDGX agent

arXiv:2605.08285v1 Announce Type: new Abstract: Autoregressive scientific forecasters often enforce physical or structural constraints by repairing each predicted state before feeding it back into the

Fast mode for Claude Opus 4.7 is now available in research preview on the API and in Claude Code.

Model ReleasesDGX agent

Anthropic has released a fast mode for Claude Opus 4.7, now available in research preview for both the API and Claude Code, offering improved performance for compatible workloads. This feature allows

Featurized Occupation Measures for Structured Global Search in Numerical Optimal Control

Model ReleasesDGX agent

arXiv:2603.16231v2 Announce Type: replace-cross Abstract: Numerical optimal control has long been split between globally structured but dimensionally intractable Hamilton--Jacobi--Bellman (HJB) method

Finer is Better (with the Right Scaling)

ResearchDGX agent

arXiv:2605.08565v1 Announce Type: new Abstract: Microscaling is a critical technique for preserving the quality of Large Language Models (LLMs) quantized to ultra-low precision formats. Intuitively, f

Fitting Is Not Enough: Smoothness in Extremely Quantized LLMs

ResearchDGX agent

arXiv:2605.08894v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance but incur high deployment costs, motivating extremely low-bit but lossy quantization. Existing

FocuSFT: Bilevel Optimization for Dilution-Aware Long-Context Fine-Tuning

AgentsDGX agent

arXiv:2605.09932v1 Announce Type: new Abstract: Large language models can now process increasingly long inputs, yet their ability to effectively use information spread across long contexts remains lim

FraudBench: A Multimodal Benchmark for Detecting AI-Generated Fraudulent Refund Evidence

Model ReleasesDGX agent

arXiv:2605.08820v1 Announce Type: cross Abstract: Artificial Intelligence (AI)-generated images have become increasingly realistic and readily adaptable to concrete real-world claims, creating new cha

From Articulated Kinematics to Routed Visual Control for Action-Conditioned Surgical Video Generation

Model ReleasesDGX agent

arXiv:2605.08712v1 Announce Type: new Abstract: Action-conditioned surgical video generation is a critical yet highly challenging problem for robotic surgery. The core difficulty is that low-dimension

From Mechanistic to Compositional Interpretability

ResearchDGX agent

arXiv:2605.08934v1 Announce Type: new Abstract: Mechanistic interpretability aims to explain neural model behaviour by reverse-engineering learned computational structure into human-understandable com

GELATO: Generative Entropy- and Lyapunov-based Adaptive Token Offloading for Device-Edge Speculative LLM Inference

Local AiDGX agent

arXiv:2605.10124v1 Announce Type: cross Abstract: The recent growth of on-device Large Language Model (LLM) inference has driven significant interest in device-edge collaborative LLM inference. As a p

Generalization Error Bounds for Picard-Type Operator Learning in Nonlinear Parabolic PDEs

Local AiDGX agent

arXiv:2605.10277v1 Announce Type: new Abstract: Operator learning for partial differential equations (PDEs) aims to learn solution operators on infinite-dimensional function spaces from finite-resolut

Global Optimization via Softmin Energy Minimization

Model ReleasesDGX agent

arXiv:2509.17815v2 Announce Type: replace Abstract: Global optimization, particularly for non-convex functions with multiple local minima, poses significant challenges for traditional gradient-based m

Hierarchical Reinforced Trader (HRT): A Bi-Level Approach for Optimizing Stock Selection and Execution

Model ReleasesDGX agent

arXiv:2410.14927v2 Announce Type: replace-cross Abstract: Automated equity trading requires converting noisy market and news signals into executable portfolio decisions under risk, turnover, and trans

How Instruction and Reasoning Data shape Post-Training: Data Quality through the Lens of Layer-wise Gradients

ResearchDGX agent

arXiv:2504.10766v2 Announce Type: replace-cross Abstract: As the post-training of large language models (LLMs) advances from instruction-following to complex reasoning tasks, understanding how differe

How LLMs Are Persuaded: A Few Attention Heads, Rerouted

SafetyDGX agent

arXiv:2605.09314v1 Announce Type: new Abstract: Language models can be persuaded to abandon factual knowledge. This vulnerability is central to AI safety, but its internal mechanism remains poorly und

How Much is Brain Data Worth for Machine Learning?

SafetyDGX agent

arXiv:2605.09243v1 Announce Type: new Abstract: If a person can solve a task, can measuring their brain make it easier to train a model to solve that task too? Recent NeuroAI work suggests that supple

Identifying Backdoored Graphs in Graph Neural Network Training: An Explanation-Based Approach with Novel Metrics

Model ReleasesDGX agent

arXiv:2403.18136v3 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) have gained popularity in numerous domains, yet they are vulnerable to backdoor attacks that can compromise their

Identifying Multi-Hit Cancer Drivers Without Massive Parallelization: A CP, MIP, and Column Generation Framework

Model ReleasesDGX agent

arXiv:2602.22551v2 Announce Type: replace-cross Abstract: Cancer is often driven by specific combinations of an estimated two to nine gene mutations, known as multi-hit combinations. Identifying these

Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards

SafetyDGX agent

arXiv:2603.00918v3 Announce Type: replace-cross Abstract: Text-to-image generation powers content creation across design, media, and data augmentation. Post-training of text-to-image generative models

Incorporating rank-free coupling and external field via an incoherent modulated spatial photonic Ising machine

ResearchDGX agent

arXiv:2512.21587v2 Announce Type: replace-cross Abstract: Spatial photonic Ising machines offer a novel optical platform for optimization and spin-model simulation, but existing diffraction-based sche

Inpainting physics: self-supervised learning for context-driven fluid simulation

Local AiDGX agent

arXiv:2605.08832v1 Announce Type: new Abstract: Neural surrogate models for computational fluid dynamics (CFD) are typically trained as forward operators that map explicit problem specifications, such

Interactive Benchmarks

TutorialsDGX agent

arXiv:2603.04737v2 Announce Type: replace Abstract: Existing reasoning evaluation paradigms suffer from different limitations: fixed benchmarks are increasingly saturated and vulnerable to contaminati

Inverse Design for Conditional Distribution Matching

ResearchDGX agent

arXiv:2605.09439v1 Announce Type: new Abstract: Generative models are powerful tools for sampling from a learned distribution P(Y mid X), and inverse-design methods invert this map to find an input x

Ister: Linear Transformer for Efficient Multivariate Time Series Forecasting

SafetyDGX agent

arXiv:2412.18798v3 Announce Type: replace-cross Abstract: Transformer-based models have achieved remarkable success in multivariate time series forecasting (MTSF) by capturing long-range dependencies.

Lattice Deduction Transformers

Model ReleasesDGX agent

arXiv:2605.08605v1 Announce Type: cross Abstract: We introduce the Lattice Deduction Transformer (LDT), a recurrent transformer that approximates logically sound deduction by projecting its latent sta

LEAF-SQL: Level-wise Exploration with Adaptive Fine-graining for Text-to-SQL Skeleton Prediction

Model ReleasesDGX agent

arXiv:2605.09295v1 Announce Type: new Abstract: Text-to-SQL translates natural language questions into executable SQL queries, enabling intuitive database access for non-experts. While large language

LITMUS: Benchmarking Behavioral Jailbreaks of LLM Agents in Real OS Environments

Model ReleasesDGX agent

arXiv:2605.10779v1 Announce Type: cross Abstract: The rapid proliferation of LLM-based autonomous agents in real operating system environments introduces a new category of safety risk beyond content s

LLM-Agnostic Semantic Representation Attack

SafetyDGX agent

arXiv:2605.08898v1 Announce Type: cross Abstract: Large Language Models (LLMs) increasingly employ alignment techniques to prevent harmful outputs. Despite these safeguards, attackers can circumvent t

LLM-FE: Automated Feature Engineering for Tabular Data with LLMs as Evolutionary Optimizers

ResearchDGX agent

arXiv:2503.14434v3 Announce Type: replace-cross Abstract: Automated feature engineering plays a critical role in improving predictive model performance for tabular learning tasks. Traditional automate

LLMs with in-context learning for Algorithmic Theoretical Physics

Model ReleasesDGX agent

arXiv:2605.08212v1 Announce Type: cross Abstract: There is an increasing number of algorithmic computations in theoretical physics. These, while conceptually simple, can nevertheless be time-consuming

Log analysis is necessary for credible evaluation of AI agents

Model ReleasesDGX agent

arXiv:2605.08545v1 Announce Type: new Abstract: Agent benchmarks typically report only final outcomes: pass or fail. This threatens evaluation credibility in three ways. First, scores may be inflated

M^2E-UAV: A Benchmark and Analysis for Onboard Motion-on-Motion Event-Based Tiny UAV Detection

Model ReleasesDGX agent

arXiv:2605.10496v1 Announce Type: new Abstract: Tiny UAV detection from an onboard event camera is difficult when the observer and target move at the same time. In this motion-on-motion regime, ego-mo

M^3: Reframing Training Measures for Discretized Physical Simulations

SafetyDGX agent

arXiv:2605.08843v1 Announce Type: new Abstract: Neural surrogate models for physical simulations are trained on discretized samples of continuous domains, where the induced empirical measure leads to

MARLaaS: Multi-Tenant Asynchronous Reinforcement Learning as a Service

SafetyDGX agent

arXiv:2605.08527v1 Announce Type: cross Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has significantly improved the reasoning capabilities of large language models (LLMs), particula

Max-pooling Network Revisited: Analyzing the Role of Semantic Probability in Multiple Instance Learning for Hallucination Detection

ResearchDGX agent

arXiv:2605.08863v1 Announce Type: new Abstract: Hallucination detection has become increasingly important for improving the reliability of large language models (LLMs). Recently, hybrid approaches suc

MDL-GBG: A Non-parametric and Interpretable Granular-Ball Generation Method for Clustering

Local AiDGX agent

arXiv:2605.08759v1 Announce Type: new Abstract: Existing granular-ball generation methods are still mainly driven by handcrafted quality measures and heuristic splitting or stopping criteria, which we

Meet physics-intern🧑‍🎓, our agentic framework for theoretical physics. It takes Gemini 3.1 Pro from 17.7% to 31.4% on CritPt, a new SOTA o…

Model ReleasesDGX agent

Meet physics-intern🧑‍🎓, our agentic framework for theoretical physics. It takes Gemini 3.1 Pro from 17.7% to 31.4% on CritPt, a new SOTA on one of the hardest benchmarks for LLMs. Theoretical physics

MemPrivacy: Privacy-Preserving Personalized Memory Management for Edge-Cloud Agents

Model ReleasesDGX agent

arXiv:2605.09530v1 Announce Type: cross Abstract: As LLM-powered agents are increasingly deployed in edge-cloud environments, personalized memory has become a key enabler of long-term adaptation and u

Meta-reinforcement learning with minimum attention

SafetyDGX agent

arXiv:2505.16741v4 Announce Type: replace Abstract: Minimum attention applies the least action principle to changes of control concerning state and time, first proposed by Brockett. The involved regul

Metal-Sci: A Scientific Compute Benchmark for Evolutionary LLM Kernel Search on Apple Silicon

Model ReleasesDGX agent

arXiv:2605.09708v1 Announce Type: cross Abstract: We present Metal-Sci, a 10-task benchmark of scientific Apple Silicon Metal compute kernels spanning six optimization regimes (stencils, all-pairs in

MicroFuse: Protein-to-Genome Expert Fusion for Microbial Operon Reasoning

Model ReleasesDGX agent

arXiv:2605.08815v1 Announce Type: new Abstract: Predicting microbial operon co-membership requires integrating two complementary biological signals: protein-scale molecular identity and genome-context

Minimizing Worst-Case Weighted Latency for Multi-Robot Persistent Monitoring: Theory and RL-Based Solutions

Model ReleasesDGX agent

arXiv:2605.09633v1 Announce Type: new Abstract: We study multi-robot persistent monitoring on weighted graphs, where node weights encode monitoring priorities and edge weights encode travel distances.

MolSight: Molecular Property Prediction with Images

Model ReleasesDGX agent

arXiv:2605.10157v1 Announce Type: cross Abstract: Every molecule ever synthesised can be drawn as a 2D skeletal diagram, yet in modern property prediction this universally available representation has

Multimodal Representation Learning Conditioned on Semantic Relations

SafetyDGX agent

arXiv:2508.17497v2 Announce Type: replace-cross Abstract: Multimodal representation learning has been largely driven by contrastive models such as CLIP, which learn a shared embedding space by alignin

MULTITEXTEDIT: Benchmarking Cross-Lingual Degradation in Text-in-Image Editing

Model ReleasesDGX agent

arXiv:2605.08163v1 Announce Type: cross Abstract: Text-in-image editing has become a key capability for visual content creation, yet existing benchmarks remain overwhelmingly English-centric and often

← Previous
1…561562563564565…1061
Next →