AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,033 results
29 Jul 2026

PEANUT: Perturbations by Eigenvector Alignment for Attacking Graph Neural Networks Under Topology-Driven Message Passing

Model ReleasesDGX agent

arXiv:2603.26136v3 Announce Type: replace Abstract: Message Passing Neural Networks (MPNNs) have achieved strong performance on tasks involving relational data. However, small perturbations to graph s

Personalization, Personas, and Forecasting in Value Alignment

Model ReleasesDGX agent

arXiv:2607.24782v1 Announce Type: new Abstract: LLM behavior may be conditioned by human identity in several ways: they may be asked to adapt to users, role-play populations, or forecast how people wo

PIcsC: Partitioning-Induced Covariate Shift Correction

Model ReleasesDGX agent

arXiv:2607.25441v1 Announce Type: new Abstract: Covariate shift across training-data partitions biases model selection and parameter estimation in cross-validation, lifelong learning, and federated le

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

PILA: Plug-and-Play Insertion for LLM-native Advertising

TutorialsDGX agent

arXiv:2607.25590v1 Announce Type: new Abstract: How to monetize large language models (LLMs) by naturally integrating sponsored content into their responses, known as LLM-native advertising, has recen

Real-time Spatial Retrieval Augmented Generation for Urban Environments

ResearchDGX agent

arXiv:2505.02271v2 Announce Type: replace Abstract: The proliferation of Generative Artificial Ingelligence (AI), especially Large Language Models, presents transformative opportunities for urban appl

REPREC: Representation Driven Parameter-Efficient Recommendation System

Model ReleasesDGX agent

arXiv:2607.24845v1 Announce Type: cross Abstract: Large language models (LLMs) have been applied to sequential recommendation by formulating it as a natural language task. Previous work has improved p

Runtime Uncertainty Monitoring for LLM-Based Multi-Agent Systems Using Bayesian Networks

AgentsDGX agent

arXiv:2607.25877v1 Announce Type: new Abstract: This paper investigates how multi-agent systems (MAS)-based on large language models (LLMs) can support actuarial risk modelling, with a particular focu

SAM-MI: A Mask-Injected Framework for Enhancing Open-Vocabulary Semantic Segmentation with SAM

Model ReleasesDGX agent

arXiv:2511.20027v2 Announce Type: replace Abstract: Open-vocabulary semantic segmentation (OVSS) aims to segment and recognize objects universally. Trained on extensive high-quality segmentation data,

SecDrift: Measuring Sector-Conditioned Security Drift in AI-Generated Code

Model ReleasesDGX agent

arXiv:2607.25225v1 Announce Type: cross Abstract: LLMs are increasingly used for code generation in critical infrastructure, yet the security effect of domain-specific prompting is understudied. We pr

Standard Transformers Achieve the Minimax Rate in Nonparametric Regression with C^{s,lambda} Targets

ResearchDGX agent

arXiv:2602.20555v2 Announce Type: replace-cross Abstract: The tremendous success of Transformer models in fields such as large language models and computer vision necessitates a rigorous theoretical i

Track-Leakage-Free Hold-Out Self-Validation for Photogrammetric Reconstruction: Protocol, Sensitivity, and Limits

Model ReleasesDGX agent

arXiv:2607.24852v1 Announce Type: new Abstract: Automated photogrammetric inspection emits metric measurements from a 3D reconstruction whose own correctness is normally unknown without an external su

28 Jul 2026

A Controlled Visual-Backbone Benchmark for Multimodal Short-Term Solar Irradiance Forecasting

Model ReleasesDGX agent

arXiv:2607.23633v1 Announce Type: cross Abstract: Sky-image irradiance studies often compare forecasting systems in which the image encoder, temporal model, fusion block, target definition, and traini

A Survey of Graph Transformers: Architectures, Theories and Applications

ResearchDGX agent

arXiv:2502.16533v3 Announce Type: replace-cross Abstract: Graph Transformers (GTs) have demonstrated a strong capability in modeling graph structures by addressing the intrinsic limitations of graph n

AI-generated Images Challenge Visual Trust in High-risk Scenarios

Model ReleasesDGX agent

arXiv:2607.22745v1 Announce Type: cross Abstract: Rapid advances in image generation are eroding the evidentiary value of visual content in settings where authenticity can affect public safety and per

Analyzing the Importance of Blank for CTC-Based Knowledge Distillation

ResearchDGX agent

arXiv:2506.01503v2 Announce Type: replace Abstract: With the rise of large pre-trained foundation models for automatic speech recognition new challenges appear. While the performance of these models i

b10158

Model ReleasesDGX agent

spec: add eagle3-v3 support for gpt-oss model (#25794) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XC

b10173

Model ReleasesDGX agent

model: Add Laguna-S-2.1 LLM_TYPE (#26233) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Lin

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.22996v1 Announce Type: cross Abstract: Large language models (LLMs) deployed in educational settings often behave as direct answerers: they disclose target concepts in the opening turn inst

Beyond Sequential Interaction: Benchmarking Parallel Execution and Coordination for GUI Agents

Model ReleasesDGX agent

arXiv:2607.22689v1 Announce Type: new Abstract: Graphical user interface (GUI) agents are systems powered by large multimodal models (LMMs). They perceive screen state and execute user instructions th

CC-AOS: Cost- and Horizon-Conditioned Amortized Backward Induction for Finite-Horizon Optimal Stopping

Model ReleasesDGX agent

arXiv:2607.22774v1 Announce Type: new Abstract: Finite-horizon optimal stopping is a central problem in early time-series classification, where a system must decide at each sequence prefix whether the

Compiler-Grounded Hierarchical Diagnosis for LLM-Based Triton Kernel Optimization

Model ReleasesDGX agent

arXiv:2607.23089v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled automated kernel generation and optimization, but most existing approaches rely on surface

Copula Based Fusion of Clinical and Genomic Machine Learning Risk Scores for Breast Cancer Risk Stratification

ResearchDGX agent

arXiv:2511.17605v2 Announce Type: replace Abstract: Clinical and gene-expression models predict breast cancer outcomes, but simple linear fusion ignores dependence between their risk scores. Using MET

Deblur-Avatar: Animatable Avatars from Motion-Blurred Monocular Videos

ApplicationsDGX agent

arXiv:2501.13335v4 Announce Type: replace Abstract: We introduce a novel framework for modeling high-fidelity, animatable 3D human avatars from motion-blurred monocular video inputs. Motion blur is pr

Dementia Etiology Diagnosis via Collaborative Meta Knowledge Enhancement

SafetyDGX agent

arXiv:2607.22770v1 Announce Type: cross Abstract: Although artificial intelligence (AI) has shown promising performance in several medical tasks, accurate dementia etiology diagnosis with AI remains c

EgoPlay: Event-Triggered Video Editing for Egocentric Streams

Model ReleasesDGX agent

arXiv:2607.24560v1 Announce Type: cross Abstract: We introduce EgoPlay, an event-triggered video-to-video editor for egocentric streams, obtained by fine-tuning a pretrained V2V diffusion transformer

Embeddings based Anomaly Detection for Cleaning Global Crop Type Reference Datasets

Local AiDGX agent

arXiv:2607.23908v1 Announce Type: new Abstract: High quality reference data remain a critical bottleneck for crop-type mapping at any spatial and temporal scale. Operational systems such as WorldCerea

FedSLIM: Privacy-Preserving Federated MDL-Based Descriptive Pattern Mining Across Data Silos

ApplicationsDGX agent

arXiv:2607.23236v1 Announce Type: cross Abstract: Federated learning has achieved considerable success for predictive modelling, yet federated descriptive analytics remains largely unexplored. Existin

FILLER: Feature Imputation via Latent Location Exploration and Retrieval

ApplicationsDGX agent

arXiv:2607.23295v1 Announce Type: cross Abstract: In real-world machine learning applications, incomplete observations create a fundamental challenge. Researchers have come up with several ideas to ad

From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search

SafetyDGX agent

arXiv:2607.24280v1 Announce Type: new Abstract: Agentic search enables large language models to solve knowledge-intensive tasks by interleaving multi-step reasoning with retrieval, yet optimizing this

Generalization bounds and sample complexity for remaining useful life prediction from complete degradation trajectories

SafetyDGX agent

arXiv:2607.23454v1 Announce Type: new Abstract: Data-driven remaining useful life (RUL) prediction requires complete degradation trajectories for training, yet such run-to-failure data are scarce and

Grounding latent algorithm routing in transformer reasoning

Model ReleasesDGX agent

arXiv:2607.24471v1 Announce Type: new Abstract: A central question in the in-context learning literature is whether transformers can organize episode-level adaptation around different inductive-bias f

Improving Adversarial Robustness of Zero-Shot CLIP with Confidence-Aware Weighting

SafetyDGX agent

arXiv:2510.02913v2 Announce Type: replace Abstract: Vision-language models such as CLIP demonstrate impressive zero-shot generalization but remain highly vulnerable to adversarial attacks. Prior adver

KANEx: Translating Kolmogorov-Arnold Networks' Interpretability to Medical Explainability

Local AiDGX agent

arXiv:2607.24730v1 Announce Type: cross Abstract: Computer vision models have become highly effective for medical applications, yet their black-box nature continues to undermine clinician trust. In cl

KAYROS: An Anytime and Exact Open-Source Solver for Duration-Minimization Time-Dependent Vehicle Routing. A Technical Report and a Case Study in Human-AI Engineering

Model ReleasesDGX agent

arXiv:2607.23116v1 Announce Type: cross Abstract: KAYROS is an open-source solver for duration-minimization time-dependent vehicle routing problems, with or without time windows (TDVRPTW, TDVRP). In t

Layering Virtual Try-On

Model ReleasesDGX agent

arXiv:2607.22924v1 Announce Type: new Abstract: In the real world, fashion is about layering: adding a jacket over a shirt, or a sequence of adding and removing layers, rather than just a single-layer

LazyMem: Retrieve Broadly, Construct Selectively for Efficient Long-Term Agent Memory

Model ReleasesDGX agent

arXiv:2607.22690v1 Announce Type: new Abstract: Long-term memory lets LLM agents reuse past interactions, but raw dialogue histories are verbose and information-sparse. Retrieving broadly improves evi

Low-Latency Turn-Taking via Context-Aware Preface Generation in a Real-World Dialogue Robot

ApplicationsDGX agent

arXiv:2607.23204v1 Announce Type: cross Abstract: Large language model (LLM)-based dialogue systems suffer response delays because generation begins only after final speech recognition. While fixed fi

Multiview Multi-Person Human Mesh Recovery Under Large Scenes with Occlusions

Model ReleasesDGX agent

arXiv:2607.24302v1 Announce Type: new Abstract: Human mesh recovery (HMR) aims to recover 3D human meshes from images. Most existing HMR benchmarks and methods focus on either multi-person reconstruct

Neptuna: A Comprehensive Machine Learning Framework for Benchmarking Complex Multiphase Flows

Model ReleasesDGX agent

arXiv:2607.22280v2 Announce Type: replace-cross Abstract: Compressible multiphase flows involving shocks and material interfaces arise in applications such as bubble collapse and droplet breakup, wher

Offline-Online Curriculum RL for Multimodal Reasoning

TutorialsDGX agent

arXiv:2607.23700v1 Announce Type: new Abstract: Multimodal large language models exhibit capabilities on reasoning tasks, yet often produce flawed intermediate steps while yielding correct final answe

OrchNAS: Orchestrated Neural Architecture Search Service for Personalised Federated Edge Intelligence

Model ReleasesDGX agent

arXiv:2607.22805v1 Announce Type: cross Abstract: We propose OrchNAS, an energy-aware, personalised, federated edge intelligence framework that leverages a Neural Architecture Search Service to automa

OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows

Model ReleasesDGX agent

arXiv:2510.24411v3 Announce Type: replace Abstract: Computer-using agents powered by Vision-Language Models (VLMs) have demonstrated human-like capabilities in operating digital environments like mobi

PRISM: Prompt Refinement via Image-grounded Self-rewarding Mechanism for Text-to-Image Generation

SafetyDGX agent

arXiv:2607.24353v1 Announce Type: new Abstract: Text-to-image generation models can synthesize high-quality images from natural language descriptions, but their performance remains highly sensitive to

ProvenanceGuard: Source-Aware Factuality Verification for MCP-Based LLM Agents

Model ReleasesDGX agent

arXiv:2606.18037v2 Announce Type: replace Abstract: Tool-using LLM agents increasingly use the Model Context Protocol (MCP) to answer from heterogeneous evidence sources, including search, APIs, datab

proxymate: Diagnosis and Adjustment of Proxy Estimates for Reliable Inference

Model ReleasesDGX agent

arXiv:2607.24401v1 Announce Type: cross Abstract: Proxy outcomes (such as short-term behavioral signals, model predictions, or surrogate endpoints) are frequently used in place of primary outcomes tha

SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes

Model ReleasesDGX agent

arXiv:2412.20541v2 Announce Type: replace Abstract: Memes act as cryptic tools for sharing sensitive ideas, often requiring contextual knowledge to interpret them correctly. It makes multimodal meme m

SCAIR: Schema-Conditioned Agentic Iterative Reasoning for Enterprise Knowledge Graphs

Model ReleasesDGX agent

arXiv:2607.22571v1 Announce Type: new Abstract: Knowledge Graph-based Retrieval-Augmented Generation (KG-RAG) enables natural language interaction with structured enterprise knowledge, yet existing ag

SeekJudge: A Practical Reward Framework for Reinforcement Learning in Computer-Use Agents

AgentsDGX agent

arXiv:2607.23263v1 Announce Type: new Abstract: Deciding whether a trajectory actually fulfills its instruction governs how we measure computer-use agents on long-horizon graphical-user-interface task

Spatula: Exploring On-Demand In-Situ Interfaces and Interaction for Attribute Control

Model ReleasesDGX agent

arXiv:2607.10405v2 Announce Type: replace-cross Abstract: Controlling attributes is a critical step toward achieving the final creative outcome, yet current approaches fall short in supporting users i

Structure over Depth: A Single-Block Spatio-Temporal Transformer for Multi-Entity Reasoning

TutorialsDGX agent

arXiv:2607.23077v1 Announce Type: new Abstract: Modeling multi-entity temporal data requires capturing dependencies across entities, time, and their interactions. Transformer-based approaches perform

The Cost of Knowing: A Resource-Aware Protocol for Benchmarking Hallucination Beyond Static Leaderboards

Model ReleasesDGX agent

arXiv:2607.24063v1 Announce Type: new Abstract: On standard factuality tasks, frontier models now cluster near the top of the scale. The question is therefore shifting from how factual a system is tow

The Tokenizer Tax: Quantifying and Explaining the Cross-Lingual Cost of Subword Tokenization for Indian Languages

Model ReleasesDGX agent

arXiv:2607.24276v1 Announce Type: cross Abstract: Large language models (LLMs) process text through subword tokenizers rather than directly reading characters or words. Because these tokenizers are tr

The Visual Bottleneck: Sparse-Frame Adaptation of MLLMs for Joint Spatial-Temporal Video Grounding

Local AiDGX agent

arXiv:2607.24570v1 Announce Type: cross Abstract: Large-scale video platforms process millions of uploads hourly, requiring moderation systems that can localize when and where policy violations occur

Update your chat template for dsv4 if you're using llama.cpp

Model ReleasesDGX agent

Following some recent commits in llama.cpp, preserve_thinking behavior for chat templates included in older DSV4 ggufs got broken. This makes the model pretty dumb in a coding agent context. Adding kw

Visual Token Compression Enhances Robustness of MLLMs

SafetyDGX agent

arXiv:2607.22716v1 Announce Type: new Abstract: In this paper, we show for the first time that visual token pruning enhances the robustness of Multimodal Large Language Models (MLLMs), mitigating vuln

WaveZip: Wavelet-Driven Space-Time Decoupling for Video Token Condensation

ResearchDGX agent

arXiv:2607.23265v1 Announce Type: new Abstract: Existing Large Vision-Language Models (LVLMs) struggle with long-form video understanding due to the quadratic computational cost of visual tokens. Whil

xMIx: High-Performance Serving-Time Platform for Mechanistic Interpretability Apps

SafetyDGX agent

arXiv:2607.22595v1 Announce Type: new Abstract: Mechanistic interpretability (MI) has emerged as a powerful approach for analyzing and intervening in inference computations, with a growing number of a

27 Jul 2026

An opinionated guide to which AI to use to do stuff

Model ReleasesDGX agent

An opinionated guide to which AI to use to do stuff It's interesting watching the evolution of Ethan Mollick's guide over time. A year ago it was still all about chat - ChatGPT, Claude, Gemini - with

b10155

Model ReleasesDGX agent

mtmd: support MiMo-V2.5 audio input (RVQ-based model) (#26190) gguf converter for mimo audio fix conv cpp impl nits nits 2 Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple

Dynamic Commonsense Coordination for Empathetic Response Generation

Model ReleasesDGX agent

arXiv:2607.22136v1 Announce Type: new Abstract: Empathetic Response Generation (ERG) requires models to recognize users' emotions and generate empathetic responses. Commonsense knowledge has been show

← Previous
1…383384385386387…1051
Next →