AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,602 results
5 May 2026

HiFi-Mamba: Dual-Stream W-Laplacian Enhanced Mamba for High-Fidelity MRI Reconstruction

ResearchDGX agent

arXiv:2508.09179v3 Announce Type: replace-cross Abstract: Reconstructing high-fidelity MR images from undersampled k-space data remains a challenging problem in MRI. While Mamba variants for vision ta

I asked ChatGPT and Gemini to do this weird perspective portrait with my face. I gave it a front face picture and a profile, including the example artwork. That was the result.

Model ReleasesDGX agent

This post documents a user's experiment comparing ChatGPT and Gemini's ability to create perspective portrait artwork using two reference images (a front-facing photo and a profile photo) along with a

Instance-Aware Parameter Configuration in Bilevel Late Acceptance Hill Climbing for the Electric Capacitated Vehicle Routing Problem

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2605.00572v1 Announce Type: new Abstract: Algorithm performance in combinatorial optimization is highly sensitive to parameter settings, while a single globally tuned configuration often fails t

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation

Model ReleasesDGX agent

arXiv:2512.21788v3 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning of Diffusion Transformers (DiTs) for diverse, multi-conditional tasks often suffers from task interference when usin

Interpretable Difficulty-Aware Knowledge Tracing in Tutor-Student Dialogues

ResearchDGX agent

arXiv:2605.01097v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have led to the development of AI-powered tutoring systems that provide interactive support via dialogue

Isotropic Fourier Neural Operators

TutorialsDGX agent

arXiv:2605.02597v1 Announce Type: new Abstract: Fourier Neural Operators are deep learning models that learn mappings between function spaces and can be used to learn and solve partial differential eq

Logit-KL Flow Matching: Non-Autoregressive Text Generation via Sampling-Hybrid Inference

ResearchDGX agent

arXiv:2411.16821v5 Announce Type: replace Abstract: Non-autoregressive (NAR) language models offer notable efficiency in text generation by circumventing the sequential bottleneck of autoregressive de

MedScribe: Clinically Grounded CT Reporting through Agentic Workflows

AgentsDGX agent

arXiv:2605.01779v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown potential for automated radiology report generation, yet existing approaches rely on global embedding compressi

Minimizing Collateral Damage in Activation Steering

SafetyDGX agent

arXiv:2605.01167v1 Announce Type: new Abstract: Activation steering is a method for controlling Large Language Model (LLM) behavior by intervening in its internal representations to increase the align

MPCS: Neuroplastic Continual Learning via Multi-Component Plasticity and Topology-Aware EWC

Model ReleasesDGX agent

arXiv:2605.02509v1 Announce Type: new Abstract: Continual learning systems face a fundamental tension between plasticity -- acquiring new knowledge -- and stability -- retaining prior knowledge. We in

Multi-Perspective Transformers in ARC-AGI-2 Challenge

Model ReleasesDGX agent

arXiv:2605.01154v1 Announce Type: new Abstract: ARC-AGI-2 is a benchmark of human-intuitive visual puzzles that measures a machine's ability to generalize from limited examples, interpret symbolic mea

New ways to buy ChatGPT ads

Model ReleasesDGX agent

OpenAI introduced new advertising purchasing options for ChatGPT, expanding how businesses can buy ad placements within the platform. These new methods likely provide advertisers with additional flexi

Not quietly. Loudly! 📣📣📣 deepagents-cli🔥

AgentsDGX agent

Not quietly. Loudly! 📣📣📣 deepagents-cli🔥 deepagents-cli is quietly becoming the best place to start coding with open weight models. we've been investing heavily in making it a harness that's truly mod

OmniEncoder: See, Hear, and Feel Continuous Motion Like Humans With One Encoder

ResearchDGX agent

arXiv:2605.01506v1 Announce Type: new Abstract: Recent advances in omni-modal large language models have enabled remarkable progress in joint vision-audio understanding. However, prevailing architectu

Parameter Space Analysis through Guided Visual Interpolations

Model ReleasesDGX agent

arXiv:2509.19202v2 Announce Type: replace-cross Abstract: We propose Parameter Space Analysis through Guided Visual Interpolations (ParamInter), a novel tool for high-dimensional input parameter space

Perturb and Correct: Post-Hoc Ensembles using Affine Redundancy

ResearchDGX agent

arXiv:2605.01632v1 Announce Type: new Abstract: Models that are indistinguishable on in-distribution data can behave very differently under distribution shift. We introduce Perturb-and-Correct (P&C),

PhaseNet++: Phase-Aware Frequency-Domain Anomaly Detection for Industrial Control Systems via Phase Coherence Graphs

Model ReleasesDGX agent

arXiv:2605.00929v1 Announce Type: new Abstract: Multivariate time series anomaly detection in ICS has attracted growing attention due to the increasing threat of cyber-physical attacks on critical inf

Pixel-to-4D: Camera-Controlled Image-to-Video Generation with Dynamic 3D Gaussians

ApplicationsDGX agent

arXiv:2601.00678v2 Announce Type: replace Abstract: Humans excel at forecasting the future dynamics of a scene given just a single image. Video generation models that can mimic this ability are an ess

Pretraining on Sleep Data Improves non-Sleep Biosignal Tasks

ResearchDGX agent

arXiv:2605.02500v1 Announce Type: new Abstract: Sleep foundation models have recently demonstrated strong performance on in-domain polysomnography tasks, including sleep staging, apnea detection, and

Psychologically Potent, Computationally Invisible: LLMs Generate Social-Comparison Triggers They Fail to Detect

Model ReleasesDGX agent

arXiv:2605.01017v1 Announce Type: new Abstract: We introduce Xiaohongshu Social Comparison Reader Elicitation (XHS-SCoRE), a reader-grounded benchmark for detecting if a text-only Xiaohongshu (RedNote

Public sector momentum and mission impact at Google Cloud Next ‘26

Model ReleasesDGX agent

The agentic era is here, and the public sector is at the forefront of leading this transformation.At Google Cloud Next ’26, it was clear that academia and public sector organizations are no longer jus

Quasi-Static Control of Discrete Cosserat Rod

ResearchDGX agent

arXiv:2605.01395v1 Announce Type: cross Abstract: In this paper, we design feedback control laws for soft robots modelled using the Cosserat rod, which is spatially discretised using the Piecewise Con

RAFNet: Region-Aware Fusion Network for Pansharpening

Model ReleasesDGX agent

arXiv:2605.02184v1 Announce Type: new Abstract: Pansharpening aims to generate high-resolution multispectral (HRMS) images by fusing low-resolution multispectral (LRMS) and high-resolution panchromati

Scaling Sequence-to-Sequence Generative Neural Rendering

ResearchDGX agent

arXiv:2510.04236v3 Announce Type: replace Abstract: We present Kaleido, a family of generative models designed for photorealistic, unified object- and scene-level neural rendering. Kaleido operates on

SRTJ: Self-Evolving Rule-Driven Training-Free LLM Jailbreaking

Model ReleasesDGX agent

arXiv:2605.00974v1 Announce Type: cross Abstract: LLMs are increasingly equipped with safety alignment mechanisms, yet recent studies demonstrate that they remain vulnerable to jailbreaking attacks th

To Call or Not to Call: A Framework to Assess and Optimize LLM Tool Calling

AgentsDGX agent

arXiv:2605.00737v1 Announce Type: new Abstract: Agentic AI architectures augment LLMs with external tools, unlocking strong capabilities. However, tool use is not always beneficial; some calls may be

Toward Culturally Grounded Natural Language Processing

Model ReleasesDGX agent

arXiv:2603.26013v2 Announce Type: replace Abstract: Multilingual NLP is often treated as a route to global inclusion, but linguistic coverage and cultural competence frequently diverge. This paper syn

Training-Free Time Series Classification via In-Context Reasoning with LLM Agents

Model ReleasesDGX agent

arXiv:2510.05950v2 Announce Type: replace Abstract: Time series classification (TSC) spans diverse application scenarios, yet labeled data are often scarce, making task-specific training costly and in

TrajRAG: Retrieving Geometric-Semantic Experience for Zero-Shot Object Navigation

TutorialsDGX agent

arXiv:2605.01700v1 Announce Type: new Abstract: Existing zero-shot Object Goal Navigation (ObjectNav) methods often exploit commonsense knowledge from large language or vision-language models to guide

Trust, but Verify: Peeling Low-Bit Transformer Networks for Training Monitoring

Local AiDGX agent

arXiv:2605.02853v1 Announce Type: new Abstract: Understanding whether deep neural networks are effectively optimized remains challenging, as training occurs in highly nonconvex landscapes and standard

Two-Pass Zero-Shot Temporal-Spatial Grounding of Rare Traffic Events in Surveillance Video

Model ReleasesDGX agent

arXiv:2605.01512v1 Announce Type: new Abstract: Grounding traffic accidents in real CCTV footage is a rare-event problem where training on labeled accident video is often prohibited, yet accurate join

Unlocking large scale AI training networks with MRC (Multipath Reliable Connection)

Model ReleasesDGX agent

MRC (Multipath Reliable Connection) is OpenAI's networking technology designed to enable efficient large-scale AI training by improving communication reliability and performance across distributed sup

Variational Matrix-Learning Fourier Networks for Parametric Multiphysics Surrogates

Model ReleasesDGX agent

arXiv:2605.02280v1 Announce Type: new Abstract: Multiphysics simulation is critical for system-technology co-optimization (STCO) in chiplet-based design, but repeated finite-element solutions of PDE-g

VAUQ: Vision-Aware Uncertainty Quantification for LVLM Self-Evaluation

ApplicationsDGX agent

arXiv:2602.21054v2 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) frequently hallucinate, limiting their safe deployment in real-world applications. Existing LLM self-eval

VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning

Local AiDGX agent

arXiv:2601.20055v2 Announce Type: replace Abstract: Despite the syntactic fluency of Large Language Models (LLMs), ensuring their logical correctness in high-stakes domains remains a fundamental chall

ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking

ResearchDGX agent

arXiv:2605.02638v1 Announce Type: new Abstract: Cross-view Referring Multi-Object Tracking (CRMOT) aims to track multiple objects specified by natural language across multiple camera views, with globa

Watermarking LLM Agent Trajectories

Model ReleasesDGX agent

arXiv:2602.18700v2 Announce Type: replace-cross Abstract: LLM agents rely heavily on high-quality trajectory data to guide their problem-solving behaviors, yet producing such data requires substantial

When Good OCR Is Not Enough: Benchmarking OCR Robustness for Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2605.00911v1 Announce Type: new Abstract: Industrial Retrieval-Augmented Generation (RAG) systems depend on optical character recognition (OCR) to transform visual documents into text. Existing

When Less Is More: Simplicity Beats Complexity for Physics-Constrained InSAR Phase Unwrapping

Model ReleasesDGX agent

arXiv:2605.00896v1 Announce Type: new Abstract: Operational phase unwrapping is the primary computational bottleneck in InSAR-based volcanic and seismic monitoring. We challenge the industry trend of

WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization

HardwareDGX agent

arXiv:2605.02262v1 Announce Type: cross Abstract: Recently, video language models (VLMs) have been applied in various fields. However, the visual token sequence of the VLM is too long, which may cause

XekRung Technical Report

TutorialsDGX agent

arXiv:2605.00072v1 Announce Type: cross Abstract: We present XekRung, a frontier large language model for cybersecurity, designed to provide comprehensive security capabilities. To achieve this, we de

4 May 2026

A Comparative Study of QSPR Methods on a Unique Multitask PAMPA dataset

ResearchDGX agent

arXiv:2605.00508v1 Announce Type: new Abstract: We present a unique, multitask dataset comprising 143 drug and drug candidate molecules, each evaluated on in vitro, parallel artificial-membrane permea

A Novel Patch-Based TDA Approach for Computed Tomography Imaging

ResearchDGX agent

arXiv:2512.12108v5 Announce Type: replace Abstract: The development of machine learning models based on computed tomography (CT) imaging has been a major focus due to the promise that imaging holds fo

April 2026 newsletter

Model ReleasesDGX agent

I just sent out the April edition of my sponsors-only monthly newsletter. If you are a sponsor (or if you start a sponsorship now) you can access it here. In this month's newsletter: Opus 4.7 and GPT-

claudely: launch Claude Code against Local LLM provider like LM Studio / Ollama / llama.cpp without trashing your real claude config

Model ReleasesDGX agent

claudely is a tool that enables users to run Claude Code against local LLM providers such as LM Studio, Ollama, or llama.cpp while preserving their existing Claude configuration. The tool allows devel

Comparing Exploration-Exploitation Strategies of LLMs and Humans: Insights from Standard Multi-armed Bandit Experiments

ResearchDGX agent

arXiv:2505.09901v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate or automate human behavior in complex sequential decision-making settings. A na

CompleteRXN: Toward Completing Open Chemical Reaction Databases

Model ReleasesDGX agent

arXiv:2605.00222v1 Announce Type: new Abstract: Chemical reaction datasets such as USPTO suffer from substantial incompleteness, frequently missing byproducts, co-reactants, and stoichiometric coeffic

CURE-OOD: Benchmarking Out-of-Distribution Detection for Survival Prediction

Model ReleasesDGX agent

arXiv:2605.00350v1 Announce Type: new Abstract: ``How long can I live and remain free of cancer?'' is often the first question a patient asks after receiving a cancer diagnosis and treatment. Accurate

Dynamics-Encoded Deep Learning for Robust System Identification and Parameter Estimation

Model ReleasesDGX agent

arXiv:2410.04299v2 Announce Type: replace Abstract: Incorporating a priori physics knowledge into machine learning leads to more robust and interpretable algorithms. In this work, we combine deep lear

Escaping Mode Collapse in LLM Generation via Geometric Regulation

ResearchDGX agent

arXiv:2605.00435v1 Announce Type: new Abstract: Mode collapse is a persistent challenge in generative modeling and appears in autoregressive text generation as behaviors ranging from explicit looping

Estimating LLM Grading Ability and Response Difficulty in Automatic Short Answer Grading via Item Response Theory

SafetyDGX agent

arXiv:2605.00238v1 Announce Type: new Abstract: Automated short answer grading (ASAG) with large language models (LLMs) is commonly evaluated with aggregate metrics such as macro-F1 and Cohen's kappa.

From Birdsong to Rumbles: Classifying Elephant Calls with Out-of-Species Embeddings

Local AiDGX agent

arXiv:2605.00225v1 Announce Type: cross Abstract: We show that pretrained acoustic embeddings classify elephant vocalisations at a level approaching that of end-to-end supervised neural networks, with

High-Probability Convergence in Decentralized Stochastic Optimization with Gradient Tracking

Model ReleasesDGX agent

arXiv:2605.00281v1 Announce Type: new Abstract: We study high-probability (HP) convergence guarantees in decentralized stochastic optimization, where multiple agents collaborate to jointly train a mod

LambdaRankIC: Directly Optimizing Rank IC for Financial Prediction

ApplicationsDGX agent

arXiv:2605.00501v1 Announce Type: new Abstract: In financial predictions, the performance of machine learning models is often assessed by Rank IC, which is the Spearman rank correlation between the mo

Learning Locally, Revising Globally: Global Reviser for Federated Learning with Noisy Labels

Model ReleasesDGX agent

arXiv:2412.00452v2 Announce Type: replace-cross Abstract: Conventioanl federated learning (FL) heavily depends on high-quality labels, which are often impractical in the real world, leading to the fed

M-CaStLe: Uncovering Local Causal Structures in Multivariate Space-Time Gridded Data

Model ReleasesDGX agent

arXiv:2605.00398v1 Announce Type: new Abstract: Causal graph discovery for space-time systems is challenging in high-dimensional gridded data, which often has many more grid cells than temporal observ

MoDAl: Self-Supervised Neural Modality Discovery via Decorrelation for Speech Neuroprosthesis

Model ReleasesDGX agent

arXiv:2605.00025v1 Announce Type: cross Abstract: Speech neuroprosthesis systems decode intended speech from neural activity in the absence of audible output, offering a path to restoring communicatio

Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning

ResearchDGX agent

arXiv:2605.00347v1 Announce Type: cross Abstract: Given the rapidly growing capabilities of vision-language models (VLMs), extending them to interactive decision-making tasks such as video games has e

On the Expressive Power of Contextual Relations in Transformers

ResearchDGX agent

arXiv:2603.25860v2 Announce Type: replace-cross Abstract: Transformer architectures have achieved remarkable empirical success in modeling contextual relations, yet a clear understanding of their expr

OTSS: Output-Targeted Soft Segmentation for Contextual Decision-Weight Learning

Model ReleasesDGX agent

arXiv:2605.00193v1 Announce Type: new Abstract: Many machine learning systems make constrained decisions by optimizing factorized objectives, but the context-specific objective is often treated as fix

← Previous
1…567568569570571…1061
Next →