AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
5 May 2026

Mitigating Multimodal LLMs Hallucinations via Relevance Propagation at Inference Time

ResearchDGX agent

arXiv:2605.01766v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have revolutionized the landscape of AI, demonstrating impressive capabilities in tackling complex vision and

MOSAIC: Multi-agent Orchestration for Task-Intelligent Scientific Coding

Model ReleasesDGX agent

arXiv:2510.08804v3 Announce Type: replace Abstract: We present MOSAIC, a multi-agent Large Language Model (LLM) framework for solving challenging scientific coding tasks. Unlike general-purpose coding

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.01687v1 Announce Type: new Abstract: We present MultiBreak, a scalable and diverse multi-turn jailbreak benchmark to evaluate large language model (LLM) safety. Multi-turn jailbreaks mimic

Pandora's Regret: A Proper Scoring Rule for Evaluating Sequential Search

Model ReleasesDGX agent

arXiv:2605.01936v1 Announce Type: new Abstract: In sequential search, alternatives are tested until the true class is found. Standard proper scoring rules like log loss are local, ignoring the ranking

ParaRNN: An Interpretable and Parallelizable Recurrent Neural Network for Time-Dependent Data

ResearchDGX agent

arXiv:2605.02692v1 Announce Type: cross Abstract: The proliferation of large-scale and structurally complex data has spurred the integration of machine learning methods into statistical modeling. Recu

Physics-Informed Neural Learning for State Reconstruction and Parameter Identification in Coupled Greenhouse Climate Dynamics

Model ReleasesDGX agent

arXiv:2605.02524v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have recently emerged as a promising framework for integrating data-driven learning with physical knowledge. In

Prescriptive Scaling Laws for Data Constrained Training

Model ReleasesDGX agent

arXiv:2605.01640v1 Announce Type: cross Abstract: Training compute is increasingly outpacing the availability of high-quality data. This shifts the central challenge from optimal compute allocation to

Probe-Geometry Alignment: Erasing the Cross-Sequence Memorization Signature Below Chance

Model ReleasesDGX agent

arXiv:2605.01699v1 Announce Type: new Abstract: Recent attacks show that behavioural unlearning of large language models leaves internal traces recoverable by adversarial probes. We characterise where

Projection-Free Transformers via Gaussian Kernel Attention

Model ReleasesDGX agent

arXiv:2605.02144v1 Announce Type: new Abstract: Self-attention in Transformers is typically implemented as softmax(QK^op/sqrt{d})V, where Q=XW_Q, K=XW_K, and V=XW_V are learned linear projections of t

Reinforcement Learning for LLM-based Multi-Agent Systems through Orchestration Traces

Model ReleasesDGX agent

arXiv:2605.02801v1 Announce Type: new Abstract: As large language model (LLM) agents evolve from isolated tool users into coordinated teams, reinforcement learning (RL) must optimize not only individu

Rethinking Multi-Label Node Classification: Do Tuned Classic GNNs Suffice?

Model ReleasesDGX agent

arXiv:2605.01403v1 Announce Type: new Abstract: Multi-label node classification (MLNC) has recently been addressed by increasingly complex label-aware designs that explicitly model node-label interact

Rhamba: Region-Aware Hybrid Attention-Mamba Framework for Self-Supervised Learning in Resting-State fMRI

ResearchDGX agent

arXiv:2605.01240v1 Announce Type: new Abstract: Self-supervised pretraining is promising for large-scale neuroimaging, yet the impact of region-aware masking and hybrid sequence modeling remains under

S^3-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data

AgentsDGX agent

arXiv:2605.01248v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has enabled newer capabilities in models, such as agentic tool-use for search. However, these models struggle

SAMamba3D: adapting Segment Anything for generalizable 3D segmentation of multiphase pore-scale images

Model ReleasesDGX agent

arXiv:2605.00916v1 Announce Type: new Abstract: Reliable segmentation of multiphase pore-scale X-ray images of rocks is necessary to quantify fluid saturation, connectivity, and interfacial geometry.

SemEval-2026 Task 7: Everyday Knowledge Across Diverse Languages and Cultures

Model ReleasesDGX agent

arXiv:2605.02601v1 Announce Type: new Abstract: We present our shared task on evaluating the adaptability of LLMs and NLP systems across multiple languages and cultures. The task data consist of an ex

Skills as Verifiable Artifacts: A Trust Schema and a Biconditional Correctness Criterion for Human-in-the-Loop Agent Runtimes

SafetyDGX agent

arXiv:2605.00424v1 Announce Type: cross Abstract: Agent skills -- structured packages of instructions, scripts, and references that augment a large language model (LLM) without modifying the model its

STEP: Warm-Started Visuomotor Policies with Spatiotemporal Consistency Prediction

Model ReleasesDGX agent

arXiv:2602.08245v2 Announce Type: replace Abstract: Diffusion policies have recently emerged as a powerful paradigm for visuomotor control in robotic manipulation due to their ability to model the dis

SwiftChannel: Algorithm-Hardware Co-Design for Deep Learning-Based 5G Channel Estimation

Model ReleasesDGX agent

arXiv:2605.01931v1 Announce Type: cross Abstract: Channel estimation is crucial in 5G communication networks for optimizing transmission parameters and ensuring reliable, high-speed communication. How

Task-Driven Subspace Decomposition for Knowledge Sharing and Isolation in LoRA-based Continual Learning

Model ReleasesDGX agent

arXiv:2603.00191v2 Announce Type: replace-cross Abstract: Continual Learning (CL) requires models to sequentially adapt to new tasks without forgetting old knowledge. Recently, Low-Rank Adaptation (Lo

Teaching LLMs Brazilian Healthcare: Injecting Knowledge from Official Clinical Guidelines

Model ReleasesDGX agent

arXiv:2605.01077v1 Announce Type: new Abstract: Brazil's Unified Health System (SUS) relies on official clinical guidelines that define diagnostic criteria, treatments, dosages, and monitoring procedu

The Good, the Bad, and the Sampled: a No-Regret Approach to Safe Online Classification

Model ReleasesDGX agent

arXiv:2510.01020v2 Announce Type: replace Abstract: We study sequential testing for a binary disease outcome when risk follows an unknown logistic model. At each round, the decision maker may either p

Towards Lightest Low-Light Image Enhancement Architecture for Mobile Devices

Model ReleasesDGX agent

arXiv:2507.04277v2 Announce Type: replace Abstract: Real-time low-light image enhancement on mobile and embedded devices requires models that balance visual quality and computational efficiency. Exist

VILAS: A VLA-Integrated Low-cost Architecture with Soft Grasping for Robotic Manipulation

Model ReleasesDGX agent

arXiv:2605.02037v1 Announce Type: new Abstract: We present VILAS, a fully low-cost, modular robotic manipulation platform designed to support end-to-end vision-language-action (VLA) policy learning an

4 May 2026

Borrowed Geometry: Computational Reuse of Frozen Text-Pretrained Transformer Weights Across Modalities

Model ReleasesDGX agent

arXiv:2605.00333v1 Announce Type: cross Abstract: Frozen Gemma 4 31B weights pretrained exclusively on text tokens, unmodified, transfer across modality boundaries through a thin trainable interface.

CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection

Model ReleasesDGX agent

arXiv:2605.00630v1 Announce Type: new Abstract: The proliferation of advanced AI video synthesis techniques poses an unprecedented challenge to digital video authenticity. Existing AI-generated video

Confidence Estimation in Automatic Short Answer Grading with LLMs

ResearchDGX agent

arXiv:2605.00200v1 Announce Type: new Abstract: Automatic Short Answer Grading (ASAG) with generative large language models (LLMs) has recently demonstrated strong performance without task-specific fi

Deepfakes: we need to re-think the concept of 'real' images

Model ReleasesDGX agent

arXiv:2509.21864v2 Announce Type: replace Abstract: The wide availability and low usability barrier of modern image generation models has triggered the reasonable fear of criminal misconduct and negat

End-to-End Autoregressive Image Generation with 1D Semantic Tokenizer

ResearchDGX agent

arXiv:2605.00503v1 Announce Type: new Abstract: Autoregressive image modeling relies on visual tokenizers to compress images into compact latent representations. We design an end-to-end training pipel

FedACT: Concurrent Federated Intelligence across Heterogeneous Data Sources

Model ReleasesDGX agent

arXiv:2605.00011v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative intelligence across decentralized data source devices in a privacy-preserving way. While substantial resea

FinSafetyBench: Evaluating LLM Safety in Real-World Financial Scenarios

Model ReleasesDGX agent

arXiv:2605.00706v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied in financial scenarios. However, they may produce harmful outputs, including facilitating illegal

Firestore at Next '26: Unlock agentic development, search and MongoDB compatibility

Model ReleasesDGX agent

In the era of AI agents, the distance between a big idea and a working application has never been shorter. As we lean more heavily on agents to help us build applications, a critical question remains:

Flow matching for Sentinel-2 super-resolution: implementation, application, and implications

ResearchDGX agent

arXiv:2605.00367v1 Announce Type: new Abstract: Developing robust techniques for super-resolution of satellite imagery involves navigating commonly observed trade-offs between spectral fidelity and pe

FreeRet: MLLMs as Training-Free Retrievers

SafetyDGX agent

arXiv:2509.24621v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are emerging as versatile foundations for mixed-modality retrieval. Yet, they often require heavy post-hoc

How OpenAI delivers low-latency voice AI at scale

Model ReleasesDGX agent

OpenAI describes its technical approach to delivering real-time voice AI services with minimal latency across large user bases, likely covering infrastructure optimization, model serving strategies, a

Hypergraph and Latent ODE Learning for Multimodal Root Cause Localization in Microservices

Model ReleasesDGX agent

arXiv:2605.00351v1 Announce Type: new Abstract: Root cause localization in cloud native microservice systems requires modeling complex service dependencies, irregular temporal dynamics, and heterogene

Learning from the Unseen: Generative Data Augmentation for Geometric-Semantic Accident Anticipation

Model ReleasesDGX agent

arXiv:2605.00051v1 Announce Type: new Abstract: Anticipating traffic accidents is a critical yet unresolved problem for autonomous driving, hindered by the inherent complexity of modeling interactions

LLM-Oriented Information Retrieval: A Denoising-First Perspective

Model ReleasesDGX agent

arXiv:2605.00505v1 Announce Type: cross Abstract: Modern information retrieval (IR) is no longer consumed primarily by humans but increasingly by large language models (LLMs) via retrieval-augmented g

Powering the Inference Era: Inside the DigitalOcean AI-Native Cloud

IndustryDGX agent

DigitalOcean outlines its AI-native cloud infrastructure designed to support the inference phase of AI model deployment, emphasizing tools and services that enable businesses to run and scale AI model

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies

Model ReleasesDGX agent

arXiv:2412.02125v2 Announce Type: replace-cross Abstract: Goal-conditioned policies enable decision-making models to execute diverse behaviors based on specified goals, yet their downstream performanc

shipped Promptloop 🌀 Claude Code but for your prompts. a CLI agent that runs the full prompt-eval loop in one chat session: create test cas…

Model ReleasesDGX agent

shipped Promptloop 🌀 Claude Code but for your prompts. a CLI agent that runs the full prompt-eval loop in one chat session: create test cases → run across models → generate reports → propose prompt →

The distillation panic

ResearchDGX agent

The article discusses concerns about the widespread use of knowledge distillation in AI development, where larger models' outputs are used to train smaller models, potentially creating a cycle of degr

TimesNet-Gen: Deep Learning-based Site Specific Strong Motion Generation

Local AiDGX agent

arXiv:2512.04694v3 Announce Type: replace Abstract: Effective earthquake risk reduction relies on accurate site-specific evaluations, which require models capable of representing the influence of loca

Try Grok

Model ReleasesDGX agent

Try Grok Grok 4.3 just became the smartest AI in the world at law and money It took #1 on TWO brutal private tests no other model could win on “Vals AI” benchmarks #1 CaseLaw (v2) - 79.31% accuracy Pr

Understanding Cognitive States from Head & Hand Motion Data

ResearchDGX agent

arXiv:2509.24255v2 Announce Type: replace-cross Abstract: As virtual reality (VR) becomes widespread, head and hand motion data captured by consumer systems has become substantially more common. Howev

Unlearning Offline Stochastic Multi-Armed Bandits

ResearchDGX agent

arXiv:2605.00638v1 Announce Type: new Abstract: Machine unlearning aims to unlearn data points from a learned model, offering a principled way to process data-deletion requests and mitigate privacy ri

VLAs are Confined yet Capable of Generalizing to Novel Instructions

Model ReleasesDGX agent

arXiv:2505.03500v5 Announce Type: replace Abstract: Vision-language-action models (VLAs) often achieve high performance on demonstrated tasks but struggle significantly when required to extrapolate, c

1 May 2026

A Reproducibility Study of LLM-Based Query Reformulation

Model ReleasesDGX agent

arXiv:2604.27421v1 Announce Type: cross Abstract: Large Language Models (LLMs) are now widely used for query reformulation and expansion in Information Retrieval, with many studies reporting substanti

AesRM: Improving Video Aesthetics with Expert-Level Feedback

Model ReleasesDGX agent

arXiv:2604.28078v1 Announce Type: new Abstract: Despite rapid advances in photorealistic video generation, real-world applications such as filmmaking require video aesthetics, e.g., harmonious colors

BicKD: Bilateral Contrastive Knowledge Distillation

SafetyDGX agent

arXiv:2602.01265v2 Announce Type: replace Abstract: Knowledge distillation (KD) is a machine learning framework that transfers knowledge from a teacher model to a student model. The vanilla KD propose

COHERENCE: Benchmarking Fine-Grained Image-Text Alignment in Interleaved Multimodal Contexts

Model ReleasesDGX agent

arXiv:2604.27389v1 Announce Type: cross Abstract: In recent years, Multimodal Large Language Models (MLLMs) have achieved remarkable progress on a wide range of multimodal benchmarks. Despite these ad

Decoding Scientific Experimental Images: The SPUR Benchmark for Perception, Understanding, and Reasoning

Model ReleasesDGX agent

arXiv:2604.27604v1 Announce Type: new Abstract: We introduce SPUR, a comprehensive benchmark for scientific experimental image perception, understanding, and reasoning, comprising 4,264 question-answe

DeepTutor: Towards Agentic Personalized Tutoring

Model ReleasesDGX agent

arXiv:2604.26962v1 Announce Type: cross Abstract: Education represents one of the most promising real-world applications for Large Language Models (LLMs). However, conventional tutoring systems rely o

Diagnosing Capability Gaps in Fine-Tuning Data

Model ReleasesDGX agent

arXiv:2604.27547v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) for domain-specific tasks requires training datasets that comprehensively cover the target capabilities a pract

Energy-Efficient Plant Monitoring via Knowledge Distillation

ApplicationsDGX agent

arXiv:2604.27178v1 Announce Type: new Abstract: Recent advances in large-scale visual representation learning have significantly improved performance in plant species and plant disease recognition tas

ExoActor: Exocentric Video Generation as Generalizable Interactive Humanoid Control

ApplicationsDGX agent

arXiv:2604.27711v1 Announce Type: new Abstract: Humanoid control systems have made significant progress in recent years, yet modeling fluent interaction-rich behavior between a robot, its surrounding

Fidelity, Diversity, and Privacy: A Multi-Dimensional LLM Evaluation for Clinical Data Augmentation

Model ReleasesDGX agent

arXiv:2604.27014v1 Announce Type: new Abstract: The scarcity of high-quality annotated medical data, particularly in mental health, poses a significant bottleneck for training robust machine learning

FineState-Bench: Benchmarking State-Conditioned Grounding for Fine-grained GUI State Setting

Model ReleasesDGX agent

arXiv:2604.27974v1 Announce Type: new Abstract: Despite the rapid progress of large vision-language models (LVLMs), fine-grained, state-conditioned GUI interaction remains challenging. Current evaluat

Grounding Agent Memory in Contextual Intent

Model ReleasesDGX agent

arXiv:2601.10702v2 Announce Type: replace-cross Abstract: Deploying large language models in long-horizon, goal-oriented interactions remains challenging because similar entities and facts recur under

ITS-Mina: A Harris Hawks Optimization-Based All-MLP Framework with Iterative Refinement and External Attention for Multivariate Time Series Forecasting

Model ReleasesDGX agent

arXiv:2604.27981v1 Announce Type: cross Abstract: Multivariate time series forecasting plays a pivotal role in numerous real-world applications, including financial analysis, energy management, and tr

Judge, Then Drive: A Critic-Centric Vision Language Action Framework for Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.27366v1 Announce Type: new Abstract: Recent advances in vision language action (VLA) models have shown remarkable potential for autonomous driving by directly mapping multimodal inputs to c

← Previous
1…420421422423424…1059
Next →