AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
27 May 2026

The Bridge-Garden Dilemma in LLM Distillation: Why Mixing Hard and Soft Labels Works

Model ReleasesDGX agent

arXiv:2605.26246v1 Announce Type: new Abstract: Knowledge distillation (KD) transfers knowledge from a large teacher model to a smaller student. In language modeling, the student is trained either on

When Does LeJEPA Learn a World Model?

SafetyDGX agent

arXiv:2605.26379v1 Announce Type: cross Abstract: A representation that scrambles the true degrees of freedom of the world cannot support reliable planning or compositional generalization. We prove th

When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation

Model ReleasesDGX agent

arXiv:2509.26600v2 Announce Type: replace-cross Abstract: As LLMs rapidly saturate existing benchmarks, automated benchmark creation using LLMs (LLM-as-a-benchmark) -- where a model generates test inp

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
26 May 2026

Act or Clarify? Modeling Sensitivity to Uncertainty and Cost in Communication

AgentsDGX agent

arXiv:2602.02843v3 Announce Type: replace Abstract: When deciding how to act under uncertainty, agents may choose to act to reduce uncertainty or they may act despite that uncertainty. In communicativ

An Efficient Learning Method to Connect Observables

Model ReleasesDGX agent

arXiv:2503.01684v3 Announce Type: replace-cross Abstract: Constructing fast and accurate surrogate models is a key ingredient for making robust predictions in many topics. We introduce a new model, th

Anatomy-Anchored Self-Supervision: Distilling Vision Foundation Models for Invariant Ultrasound Representation

SafetyDGX agent

arXiv:2605.25402v1 Announce Type: cross Abstract: Self-supervised pre-training paradigm has gained increasing prominence for learning transferable representations in medical imaging, yet existing meth

APT-Agent: Automated Penetration Testing using Large Language Models

AgentsDGX agent

arXiv:2605.24949v1 Announce Type: cross Abstract: Penetration testing is essential to securing modern web infrastructures, yet traditional manual methods struggle to keep pace with their scale and com

Beyond Inference-Only Deployment: Comparing Weight-Based Consolidation Against Cascading Compaction

Model ReleasesDGX agent

arXiv:2605.24657v1 Announce Type: new Abstract: Major LLM platforms deploy models in an inference-only configuration: the model serves requests but never updates per-user weights. Users must repeatedl

Beyond Static Uncertainty: Modeling Temporal Uncertainty Dynamics for Probabilistic Time Series Forecasting

ApplicationsDGX agent

arXiv:2603.24254v2 Announce Type: replace-cross Abstract: Real-world time series exhibit temporally structured uncertainty: volatility clusters in turbulent regimes, dissipates in stable periods, and

Beyond the Target: From Imitation to Collaboration in Speculative Decoding

SafetyDGX agent

arXiv:2605.24793v1 Announce Type: new Abstract: Speculative decoding (SPD) accelerates large language model (LLM) inference by letting a smaller draft model propose multiple future tokens that are ver

Can Large Language Models Resolve Semantic Discrepancy in Self-Destructive Subcultures? Evidence from Jirai Kei

SafetyDGX agent

arXiv:2601.05004v2 Announce Type: replace Abstract: Self-destructive behaviors are linked to complex psychological states and can be challenging to diagnose. These behaviors may be even harder to iden

Don't Retrain, Just Reuse: Recovering Dual-Target Molecules from Single-Target Diffusion Models

ResearchDGX agent

arXiv:2605.25681v1 Announce Type: cross Abstract: Designing a single molecule that modulates two targets is a promising strategy for polypharmacology, but it remains substantially harder than standard

DRIVE: Modeling Skills at the Reasoning and Interaction Levels for Web Agents under Continual Learning

ResearchDGX agent

arXiv:2605.23939v1 Announce Type: new Abstract: Web agents require both high-level reasoning (for task decomposition) and low-level interactions (for page elements manipulation) to conduct different t

Empirical Analysis and Detection of Hallucinations in LLM-Generated Bug Report Summaries

Model ReleasesDGX agent

arXiv:2605.24137v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to generate summaries of software bug reports, including sections such as Steps-to-Reproduce (S2R),

Free the 100B Gemma 4 MoE! Gemini Flash 3.5 is out so now you can release it!

Model ReleasesDGX agent

Clem Delangue advocates for the release of a 100 billion parameter Gemma 4 Mixture of Experts model, suggesting that Gemini Flash 3.5's release creates an opportunity for this larger model to be made

Improving the Completeness and Comparability of Segment Disclosures: A Large Language Model Approach

SafetyDGX agent

arXiv:2605.23924v1 Announce Type: new Abstract: Segment-level disclosures are a central component of financial reporting, providing insight into firms' internal organization and the allocation of econ

In Search of the Ingredients of Open-Endedness: Replicating Picbreeder with Large Vision-Language Models

ApplicationsDGX agent

arXiv:2605.23908v1 Announce Type: new Abstract: We are in the midst of large-scale industrial and academic efforts to automate the processes of scientific, technological and creative production throug

Induction Meets Biology: Mechanisms of Repeat Detection in Protein Language Models

ResearchDGX agent

arXiv:2602.23179v3 Announce Type: replace Abstract: Protein sequences are abundant in repeating segments, both as exact copies and as approximate segments with mutations. These repeats are important f

Missing Pattern Recognized Diffusion Imputation Model for Missing Not At Random

ApplicationsDGX agent

arXiv:2605.25439v1 Announce Type: new Abstract: Missing data frequently arises across diverse domains, including time-series and image domains. In the real world, missing occurrences often depend on t

Mode-as-Sequence: Translating Multimodal Motion Prediction into Unified Sequential Mode Modeling

TutorialsDGX agent

arXiv:2605.24037v1 Announce Type: cross Abstract: Multimodal motion forecasting is inherently under-supervised: each training scene provides only one realized future, yet multiple plausible futures ex

Multilingual Phonological Feature Recognition with Self-Supervised Speech Models

ResearchDGX agent

arXiv:2605.25596v1 Announce Type: new Abstract: Phonological features provide a language-general and linguistically grounded representation of speech. We present PhonoQ-2.0, a multilingual frame-level

Optimal Design for Multinomial Logit Model with Applications to Best Assortment Identification

AgentsDGX agent

arXiv:2605.25592v1 Announce Type: cross Abstract: We study optimal experimental design for multinomial logit (MNL) bandits, where an agent repeatedly selects a subset of K items from a ground set of s

Palette: A Modular, Controllable, and Efficient Framework for On-demand Authorized Safety Alignment Relaxation in LLMs

Model ReleasesDGX agent

arXiv:2605.24154v1 Announce Type: new Abstract: Current safety alignment of foundation models largely follows a one-size-fits-all paradigm, applying the same refusal policy across users and contexts.

PCGRLLM: Large Language Model-Driven Reward Design for Procedural Content Generation Reinforcement Learning

AgentsDGX agent

arXiv:2502.10906v2 Announce Type: replace Abstract: Reward design plays a pivotal role in the training of game AIs, requiring substantial domain-specific knowledge and human effort. In recent years, s

Privacy-Preserving Local Language Models for Longitudinal Data Retrieval in Chronic Dermatologic Disease: Implementation in Pemphigus Patients

Local AiDGX agent

arXiv:2605.25020v1 Announce Type: new Abstract: Chronic dermatologic diseases such as pemphigus require long-term follow-up, generating extensive longitudinal clinical documentation that is difficult

SLAP: Stratified Loss-based Pruning for On-Policy Data-Efficient Instruction Tuning

Model ReleasesDGX agent

arXiv:2605.23969v1 Announce Type: new Abstract: Instruction tuning has optimized the specialized capabilities of large language models (LLMs), but it often requires extensive datasets and prolonged tr

Subspace Aggregation Query and Index Generation for Multidimensional Resource Space Model

ResearchDGX agent

arXiv:2505.02129v3 Announce Type: replace-cross Abstract: Organizing large-scale resources in a multidimensional semantic space is an approach to efficiently managing and querying resources from diffe

The LSCD Benchmark: a Testbed for Diachronic Word Meaning Tasks

Model ReleasesDGX agent

arXiv:2404.00176v3 Announce Type: replace Abstract: Lexical Semantic Change Detection (LSCD) is a complex, lemma-level task, which is usually operationalized based on two subsequently applied usage-le

TRACER: A Semantic-Aware Framework for Fine-Grained Contamination Detection in Code LLMs

Model ReleasesDGX agent

arXiv:2605.24079v1 Announce Type: cross Abstract: Data contamination is a known threat to the reliability of model evaluation. However, it remains underexplored in code large language models (LLMs), w

Weakly Supervised Camouflaged Object Detection Based on the SAM Model and Mask Guidance

TutorialsDGX agent

arXiv:2605.25385v1 Announce Type: cross Abstract: Camouflaged object detection (COD) from a single image is a challenging task due to the high similarity between objects and their surroundings. Existi

When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Agents

Model ReleasesDGX agent

arXiv:2605.24069v1 Announce Type: cross Abstract: The rise of tool-using Large Language Model (LLM) agents, standardized by protocols like the Model Context Protocol (MCP), has unlocked unprecedented

25 May 2026

ETCHR: Editing To Clarify and Harness Reasoning

Model ReleasesDGX agent

arXiv:2605.23897v1 Announce Type: cross Abstract: Multimodal Large Language Models have advanced visual reasoning, yet a purely textual chain of thought remains a bottleneck for questions that require

GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation

TutorialsDGX agent

arXiv:2506.14135v5 Announce Type: replace-cross Abstract: Accurate scene perception is critical for vision-based robotic manipulation. Existing approaches typically follow either a Vision-to-Action (V

Hinge Regression Trees and HRT-Boost: Newton-Optimized Oblique Learning for Compact Tabular Models

ApplicationsDGX agent

arXiv:2605.23422v1 Announce Type: new Abstract: Learning high-quality oblique decision trees remains a significant challenge due to the discrete and non-convex nature of split optimization. We present

Knowledge Distillation for Low-Resource Open-source Text-to-SQL Model

ApplicationsDGX agent

arXiv:2605.22843v1 Announce Type: new Abstract: Text-to-SQL converts natural language questions into executable SQL queries, enabling non-technical users to access relational databases for analytics a

PixlStash 1.3: grid loading speed, JoyCaption and bulk tag selections with your chosen model

Local AiDGX agent

PixlStash 1.3 is a Python-based image management and tagging web app that improves grid loading performance and introduces JoyCaption integration for AI-powered image captioning. The update adds suppo

When AI Takes Sides on Questions of Faith: Persistent Asymmetries in AI-Mediated Faith Guidance

ApplicationsDGX agent

arXiv:2605.22975v1 Announce Type: new Abstract: We ask whether large language models (LLMs) treat queries about religious conversion symmetrically. The answer is no. When asked for advice on hypotheti

XWind: A Cross-site Router for Large Language Model Inference Serving at Renewable Energy Farms

HardwareDGX agent

arXiv:2605.23348v1 Announce Type: cross Abstract: AI power demand is growing at an unprecedented rate while power grids are often ailing and struggle to keep up. Grid expansion comes with high capital

23 May 2026

Hyperparameter Transfer with Mixture-of-Expert Layers

Model ReleasesDGX agent

arXiv:2601.20205v3 Announce Type: replace Abstract: Mixture-of-Experts (MoE) layers have emerged as an important tool in scaling up modern neural networks by decoupling total trainable parameters from

I Built a local Stable Diffusion GUI specifically for older GPUs (GTX 1060). Features Zero-Copy ADetailer, URL Model Downloader, and real-time VRAM monitoring.

Local AiDGX agent

This post describes a custom Stable Diffusion graphical interface optimized for older graphics cards, specifically the GTX 1060, incorporating features like zero-copy ADetailer (a detail enhancement t

22 May 2026

AMEL: Accumulated Message Effects on LLM Judgments

Model ReleasesDGX agent

arXiv:2605.22714v1 Announce Type: cross Abstract: Large language models are routinely used as automated evaluators: to review code, moderate content, or score outputs, often with many items passing th

Branch-Stochastic Model Predictive Control for Motion Planning under Multi-Modal Uncertainty with Scenario Clustering

SafetyDGX agent

arXiv:2605.22600v1 Announce Type: new Abstract: Motion planning for autonomous driving must account for multi-modal uncertainty in both the intentions and trajectories of surrounding vehicles. Handlin

ConvNeXt-FD: A Fractal-Based Deep Model for Robust Biomedical Image Segmentation

ResearchDGX agent

arXiv:2605.22002v1 Announce Type: new Abstract: Biomedical image segmentation is a critical task in medical diagnosis and treatment planning, enabling precise delineation of anatomical structures and

DeepWeb-Bench: A Deep Research Benchmark Demanding Massive Cross-Source Evidence and Long-Horizon Derivation

Model ReleasesDGX agent

arXiv:2605.21482v1 Announce Type: new Abstract: Deep research, in which an agent searches the open web, collects evidence, and derives an answer through extended reasoning, is a prominent use case for

GeoWeaver: Grounding Visual Tokens with Geometric Evidence before Scene Reasoning

Model ReleasesDGX agent

arXiv:2605.22558v1 Announce Type: new Abstract: Spatio-temporal reasoning in vision-language models requires visual representations that preserve physical geometry rather than merely semantic appearan

Improved DDIM Sampling with Moment Matching Gaussian Mixtures

ResearchDGX agent

arXiv:2311.04938v5 Announce Type: replace Abstract: We propose using a Gaussian Mixture Model (GMM) as reverse transition operator (kernel) within the Denoising Diffusion Implicit Models (DDIM) framew

Intelligence per Watt: Measuring Intelligence Efficiency of Local AI

Local AiDGX agent

arXiv:2511.07885v4 Announce Type: replace-cross Abstract: Large language model (LLM) queries are predominantly processed by frontier models in centralized cloud infrastructure. Demand growth strains t

Rethinking Token Reduction for Diffusion Models via Output-Similarity-Awareness

ResearchDGX agent

arXiv:2605.22011v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) achieve superior image generation quality but suffer from quadratic computational complexity relative to token count. Whil

Terminal Constraint Model Predictive Control for Image-Based Visual Servoing of UAVs with Kalman Filter-Based Moment Loss Compensation

ResearchDGX agent

arXiv:2605.22443v1 Announce Type: new Abstract: Image-Based Visual Servoing (IBVS) provides an efficient vision-guided control paradigm for unmanned aerial vehicles (UAVs) by directly regulating image

The Blueprint: How Movix fills a gap in dental skills with specialized agentic AI

Model ReleasesDGX agent

Welcome to The Blueprint, a regular feature where we highlight how Google Cloud customers are tackling unique and common challenges across industries using the latest AI and cloud technologies. We hop

Though I will continue to insist that the path we were on would have been much clearer if o3 had been called GPT-5 and GPT-5 had been called…

Model ReleasesDGX agent

Ethan Mollick comments on OpenAI's naming convention for their AI models, specifically critiquing the decision to call the o3 model 'o3' rather than 'GPT-5,' suggesting the naming scheme creates confu

21 May 2026

AgentAtlas: Beyond Outcome Leaderboards for LLM Agents

Model ReleasesDGX agent

arXiv:2605.20530v1 Announce Type: cross Abstract: Large language model agents now act on codebases, browsers, operating systems, calendars, files, and tool ecosystems, but the benchmarks used to evalu

Bridging Structure and Language: Graph-Based Visual Reasoning for Autonomous Road Understanding

Model ReleasesDGX agent

arXiv:2605.20942v1 Announce Type: new Abstract: Structured road understanding of lane geometry, topology, and traffic element relationships is foundational to safe autonomous driving. While vision-lan

Can Vision Models Truly Forget? Mirage: Representation-Level Certification of Visual Unlearning

SafetyDGX agent

arXiv:2605.20282v1 Announce Type: new Abstract: Machine unlearning in Vertical Federated Learning (VFL) has attracted growing interest, yet existing methods certify forgetting solely using output-leve

Comparative Evaluation of Deep Learning Models for Fake Image Detection

SafetyDGX agent

arXiv:2605.20971v1 Announce Type: new Abstract: The growing sophistication of GAN-based image manipulation presents significant challenges for digital forensics. This study compares the performance of

CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning

Model ReleasesDGX agent

arXiv:2605.20247v1 Announce Type: cross Abstract: Catastrophic forgetting remains a major obstacle to continual learning in large language models (LLMs) and vision--language models (VLMs). Although Mi

CRANE: Correcting Errors in Raw Nanopore Signals Using Hidden Markov Models

ResearchDGX agent

arXiv:2603.20420v2 Announce Type: replace-cross Abstract: Nanopore sequencing can read substantially longer sequences of nucleic acid molecules, called reads, than other sequencing methods, which has

EvoStruct: Bridging Evolutionary and Structural Priors for Antibody CDR Design via Protein Language Model Adaptation

ResearchDGX agent

arXiv:2605.21485v1 Announce Type: new Abstract: Equivariant graph neural network (GNN) methods for antibody complementarity-determining region (CDR) design achieve the highest sequence recovery but su

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.20256v1 Announce Type: new Abstract: Reinforcement learning has become a cornerstone for aligning and unlocking the reasoning capabilities of large-scale models. At its core, the training l

Grok Build 0.1 might be one of the most underestimated AI models right now. We tested it in Kilo Code by asking it to build 5 websites from …

IndustryDGX agent

Grok 0.1 was evaluated on its ability to build complete websites from text prompts in a Kilo Code test, demonstrating capabilities that may exceed initial market expectations. The test involved reques

← Previous
1…259260261262263…1034
Next →