AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
26 May 2026

Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence

ResearchDGX agent

arXiv:2509.23573v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to help security analysts manage the surge of cyber threats, automating tasks from vulnerab

Unlocking Apple's Private Cloud Compute: An Analysis of Privacy-Preserving Artificial Intelligence

Local AiDGX agent

arXiv:2605.24239v1 Announce Type: cross Abstract: Many existing Artificial Intelligence (AI) solutions on mobile devices rely on an extensive collection of sensitive data, raising privacy concerns and

V3H: View Variation and View Heredity for Incomplete Multi-view Clustering

Model ReleasesDGX agent

arXiv:2011.11194v4 Announce Type: replace Abstract: Real data often appear in the form of multiple incomplete views. Incomplete multi-view clustering is an effective method to integrate these incomple

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

VineLM: Trie-Based Fine-Grained Control for Agentic Workflows

AgentsDGX agent

arXiv:2605.23914v1 Announce Type: cross Abstract: Agentic workflows interleave configurable LLM stages with tool stages and often include retries or refinement loops. Existing workflow managers profil

Volatility Surface Reconstruction using Deep Learning under No-Arbitrage Constraints

ResearchDGX agent

arXiv:2605.24031v1 Announce Type: cross Abstract: We study the reconstruction of implied volatility surfaces from sparse and noisy option quotes using deep learning models under no-arbitrage constrain

We have, as far as I can tell, no good tests of the productivity impact of the autonomous coding tools that appeared starting in December 20…

Model ReleasesDGX agent

We have, as far as I can tell, no good tests of the productivity impact of the autonomous coding tools that appeared starting in December 2025. Every paper out there is from prior to the Claude Code/C

We’ve shipped a security-guidance plugin for Claude Code that helps identify and fix vulnerabilities as you’re writing code. Available for a…

Model ReleasesDGX agent

We’ve shipped a security-guidance plugin for Claude Code that helps identify and fix vulnerabilities as you’re writing code. Available for all Claude Code users. Install from the plugin marketplace (/

When Interpretability Becomes a Liability: Adversarial Attacks on CBM Concept Layers

ResearchDGX agent

arXiv:2605.25304v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) have emerged as a cornerstone approach for interpretable machine learning, providing human-understandable intermediate

WhenLoss: Diagnosing Write and Retrieval Bottlenecks in Long-Context Memory Systems

Model ReleasesDGX agent

arXiv:2605.24579v1 Announce Type: new Abstract: Long-context memory systems often fail under fixed budgets, but end-to-end evaluation does not reveal whether evidence was discarded during compression

Word Class Representations Spontaneously Emerge from Successor Representations Trained on Natural Language

ResearchDGX agent

arXiv:2605.24585v1 Announce Type: new Abstract: Language models are typically trained to predict the next token in a sequence. Here, we explore an alternative predictive principle from reinforcement l

25 May 2026

A measurement substrate for agentic Kubernetes operations: Methodology and a case study in retrieval-compounding falsification

Model ReleasesDGX agent

arXiv:2605.23058v1 Announce Type: cross Abstract: Empirical claims about autonomous Kubernetes operations agents are largely unfalsifiable. Published work reports observational results without control

Agentic Proving for Program Verification

Model ReleasesDGX agent

arXiv:2605.23772v1 Announce Type: new Abstract: Agentic systems have recently emerged as state-of-the-art approaches for automated theorem proving in formal mathematics. To assess how far these capabi

AI Assurance: A Comprehensive Testing Strategy for Enterprise AI Systems

AgentsDGX agent

arXiv:2605.23459v1 Announce Type: cross Abstract: Enterprise AI systems, built on large language models, retrieval pipelines and autonomous agents, introduce a class of risks that traditional software

AI-Friendly LaTeX: Using LaTeX Code as a Knowledge Source for Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2605.22923v1 Announce Type: cross Abstract: Large language models can answer questions about textbooks, lecture notes, and programming exercises more reliably when their answers are grounded in

Android app Ollama Talk now on playstore

Local AiDGX agent

Ollama Talk is an Android app that connects to an Ollama server, enabling conversations with AI models like Llama and Mistral . The app features an intuitive chat interface with real-time conversation

Apparently this clip is too spicy! So let's try it this way! Examples of Director with LTX 2.3 and a few different techniques.

Local AiDGX agent

This post discusses workarounds for content restrictions in Stable Diffusion, specifically showcasing examples of using the Director feature with LTX 2.3 model and alternative techniques to generate c

ArcMark: Distortion-Free Multi-Byte LLM Watermark via Optimal Transport

ResearchDGX agent

arXiv:2602.07235v2 Announce Type: replace-cross Abstract: Watermarking is an important tool for promoting the responsible use of large language models (LLMs). Existing watermarks insert a signal into

Are Targeted Data Poisoning Attacks as Effective as We Think?

ResearchDGX agent

arXiv:2509.06896v2 Announce Type: replace Abstract: Targeted data poisoning attacks manipulate model predictions on specific test samples by injecting malicious data into training. Yet existing evalua

As X, Do Y: How Persona and Task Combine in Instruction-Tuned LLMs

Model ReleasesDGX agent

arXiv:2605.23147v1 Announce Type: cross Abstract: Role prompts of the form As X, do Y admit a clean linear decomposition at one specific site in the residual stream: the prompt-to-answer transition --

b9311

Local AiDGX agent

b9311 is a release build version of llama.cpp, an open-source C/C++ framework for large language model inference. The project enables LLM inference in C/C++ , offering optimized performance across var

BOHM: Zero-Cost Hierarchical Attribution for Compound AI Systems

Model ReleasesDGX agent

arXiv:2605.22866v1 Announce Type: new Abstract: Compound AI systems route tasks through hierarchies of specialised components. Attribution is dominated by Shapley-based methods (SHAP), which decompose

Canadians be like come to our Toronto tech week party

Model ReleasesDGX agent

Toronto Tech Week is an annual event in Toronto, Canada that brings together technology professionals, entrepreneurs, and innovators for networking, learning, and celebration of the local tech ecosyst

CHASD: Language Increment-Calibrated Contrastive Decoding against Hallucination in LVLMs

Local AiDGX agent

arXiv:2605.23344v1 Announce Type: cross Abstract: Large Vision-Language Models have shown strong multimodal reasoning capabilities, yet they remain susceptible to object hallucinations when language p

CHRONOS: Temporally-Aware Multi-Agent Coordination for Evolving Data Marketplaces

Model ReleasesDGX agent

arXiv:2605.23887v1 Announce Type: cross Abstract: Temporal knowledge-graph data marketplaces face three coupled failures in static designs: stale hybrid index shortcuts reduce recall as edges evolve,

Composing People Together: Iterative Pose-Image Generation for Multi-Person Interaction Scenes

SafetyDGX agent

arXiv:2605.23178v1 Announce Type: new Abstract: Despite recent progress, text-to-image models still struggle to generate semantically diverse and compositionally accurate multi-person interaction scen

Computable Fairness: Boltzmann-Softmax Control for AI Resource Allocation

Model ReleasesDGX agent

arXiv:2605.22827v1 Announce Type: cross Abstract: In large-scale AI systems, allocating scarce resources such as GPU compute time and bandwidth among multiple agents is a critical challenge. Conventio

Coupling-Robust Accuracy in Multiphysics Physics Informed Neural Networks via Kronecker-Preconditioned Optimization

Model ReleasesDGX agent

arXiv:2605.23391v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) for coupled multiphysics systems suffer systematic accuracy degradation as inter-equation coupling strengthens.

Deja Vu in Plots: Leveraging Cross-Session Evidence with Retrieval-Augmented LLMs for Live Streaming Risk Assessment

Local AiDGX agent

arXiv:2601.16027v2 Announce Type: replace Abstract: The rise of live streaming has transformed online interaction, enabling massive real-time engagement but also exposing platforms to complex risks su

Design and Report Benchmarks for Knowledge Work

Model ReleasesDGX agent

arXiv:2605.23262v1 Announce Type: new Abstract: The development of LLM agents has led to a growing body of work on knowledge-work AI, including coding, research, and healthcare. However, current knowl

Diffusion-based Denoising Beats Vanilla Score Matching in Parameter Estimation: A Theoretical Explanation

Model ReleasesDGX agent

arXiv:2605.22950v1 Announce Type: cross Abstract: Score matching is an alternative to maximum likelihood estimation when the normalizing constant is unknown or too costly to evaluate. However, vanilla

Efficient Gradient Estimation for Parameterized Quantum Systems with Lie Algebraic Symmetries

Model ReleasesDGX agent

arXiv:2404.05108v3 Announce Type: replace-cross Abstract: Gradient estimation is a central challenge in training parameterized quantum circuits (PQCs) for hybrid quantum-classical optimization and lea

Energy per Successful Goal: Goal-Level Energy Accounting for Agentic AI Systems

SafetyDGX agent

arXiv:2605.22883v1 Announce Type: new Abstract: Current AI energy benchmarks measure consumption at the granularity of a single model invocation or training run. For classical single-turn workloads th

Entropy Equivalence Testing

Model ReleasesDGX agent

arXiv:2605.23225v1 Announce Type: cross Abstract: We introduce the problem of entropy equivalence testing for probability distributions, a relaxation of the well-studied closeness testing problem, whe

Evaluating Memory Structure in LLM Agents

Model ReleasesDGX agent

arXiv:2602.11243v2 Announce Type: replace-cross Abstract: Modern LLM-based agents and chat assistants rely on long-term memory frameworks to store reusable knowledge, recall user preferences, and augm

Exploitation of KnowledgeDeliver via ViewState Deserialization Vulnerability

Model ReleasesDGX agent

Written by: Takahiro Sugiyama, Peter Revelant, Mathew Potaczek Introduction In late 2025, Mandiant responded to a security incident involving a compromised web server running KnowledgeDeliver. Knowled

Fast-dDrive: Efficient Block-Diffusion VLM for Autonomous Driving

SafetyDGX agent

arXiv:2605.23163v1 Announce Type: new Abstract: End-to-end autonomous driving via Vision-Language-Action (VLA) models demands a precarious balance between high-fidelity trajectory planning and efficie

FastKernels: Benchmarking GPU Kernel Generation in Production

Model ReleasesDGX agent

arXiv:2605.23215v1 Announce Type: cross Abstract: LLM-based agents for GPU kernel generation are advancing rapidly, yet their progress is fundamentally constrained by the benchmarks they optimize agai

FuRA: Full-Rank Parameter-Efficient Fine-Tuning with Spectral Preconditioning

Model ReleasesDGX agent

arXiv:2605.22869v1 Announce Type: new Abstract: Both full fine-tuning (Full FT) and parameter-efficient fine-tuning methods such as LoRA introduce weight updates without accounting for the spectral st

GEMQ: Global Expert-Level Mixed-Precision Quantization for MoE LLMs

ResearchDGX agent

arXiv:2605.23078v1 Announce Type: cross Abstract: Mixture-of-Experts Large Language Models (MoE-LLMs) achieve strong performance but incur substantial memory overhead due to massive expert parameters.

GeoMAE: Masking Representation Learning for Spatio-Temporal Graph Forecasting with Missing Values

ApplicationsDGX agent

arXiv:2508.14083v3 Announce Type: replace-cross Abstract: The ubiquity of missing data in urban intelligence systems, attributable to adverse environmental conditions and equipment failures, poses a s

Graph Alignment Topology as an Inductive Bias for Grounding Detection

SafetyDGX agent

arXiv:2605.22963v1 Announce Type: cross Abstract: Large Language Models (LLMs) are optimized to produce distributionally plausible continuations rather than to explicitly verify whether generated prop

. @GrooveJonesXR needed to deliver the impossible: giant NFL-licensed Crocs parachuting into Dick's Sporting Goods parking lots; hyper-reali…

Model ReleasesDGX agent

. @GrooveJonesXR needed to deliver the impossible: giant NFL-licensed Crocs parachuting into Dick's Sporting Goods parking lots; hyper-realistic, multi-location, vertical 9:16 on a holiday deadline. A

HawkesLLM: Semantic Uncertainty Propagation in Agentic Text Simulation

Local AiDGX agent

arXiv:2605.23043v1 Announce Type: new Abstract: Agentic text-simulation systems write in sequence, with each item becoming possible context for later steps. That makes uncertainty path-dependent: an e

Hermes Agent now can orchestrate the @OpenHandsDev agents with a new optional skill! `hermes update` then do `hermes skills install official…

Model ReleasesDGX agent

Hermes Agent now can orchestrate the @OpenHandsDev agents with a new optional skill! `hermes update` then do `hermes skills install official/autonomous-ai-agents/openhands` Reminder: You can already d

HTMuon: Improving Muon via Heavy-Tailed Spectral Correction

Model ReleasesDGX agent

arXiv:2603.10067v2 Announce Type: replace-cross Abstract: Muon has recently shown promising results in LLM training. In this work, we study how to further improve Muon. We argue that Muon's orthogonal

✅Implicit caching is now live on Qwen3.7-Max — kicks in automatically, no setup needed. ⚡️Faster + cheaper out of the box. Need higher, more…

Model ReleasesDGX agent

✅Implicit caching is now live on Qwen3.7-Max — kicks in automatically, no setup needed. ⚡️Faster + cheaper out of the box. Need higher, more deterministic hit rates? Try explicit caching instead. 🙌 🔗B

In his ~43,000-word encyclical, the Pope urged governments to slow down AI development and decried 'new forms of slavery' in AI and tech supply chains (Joshua McElwee/Reuters)

Model ReleasesDGX agent

Joshua McElwee / Reuters: In his ~43,000-word encyclical, the Pope urged governments to slow down AI development and decried “new forms of slavery” in AI and tech supply chains — Pope Leo urged govern

Inductive Deductive Synthesis: Enabling AI to Generate Formally Verified Systems

Model ReleasesDGX agent

arXiv:2605.23109v1 Announce Type: new Abstract: AI agents increasingly excel at generating, testing, and refining code. However, they fall short on tasks requiring formal guarantees of full coverage t

Instrumentation for Imitation Learning: Enhancing Training Datasets for Clothes Hanger Insertion

SafetyDGX agent

arXiv:2605.23847v1 Announce Type: new Abstract: Large behaviour models have transformed the field of robotic manipulation, but prohibitive data requirements have thus far prevented a revolution simila

IntentScore: Intent-Conditioned Action Evaluation for Computer-Use Agents

SafetyDGX agent

arXiv:2604.05157v2 Announce Type: replace Abstract: Computer-Use Agents (CUAs) leverage large language models to execute GUI operations on desktop environments, yet they generate actions without evalu

Is a Document Educational or Just Wikipedia-Style? -- Pitfalls of Classifier-Based Quality Filtering

ApplicationsDGX agent

arXiv:2605.23721v1 Announce Type: new Abstract: Classifier-based Quality Filtering has recently emerged as a fundamental technique in constructing pre-training corpora. The ability to deploy a single

Just found that if you scroll down in the Claude Code app on iPhone… Clawd starts jumping and walking around looking for apps to build

Model ReleasesDGX agent

The Claude Code app for iPhone features an interactive Easter egg where scrolling down triggers an animated character named 'Clawd' that jumps and walks around the screen, appearing to search for apps

Learning Safely Without Knowing the World:COMPASS-Hedge

Model ReleasesDGX agent

arXiv:2603.22348v3 Announce Type: replace Abstract: Online learning algorithms often face a fundamental trilemma: balancing regret guarantees between adversarial and stochastic settings and providing

MAS-Orchestra: Understanding and Improving Multi-Agent Reasoning Through Holistic Orchestration and Controlled Benchmarks

Model ReleasesDGX agent

arXiv:2601.14652v5 Announce Type: replace Abstract: While multi-agent systems (MAS) promise elevated intelligence through coordination of agents, current approaches to automatic MAS design under-deliv

Metadata Predictability Is Not Evidence Dependence: An Intervention-Based Audit for Weak-Label Benchmarks

Model ReleasesDGX agent

arXiv:2605.23701v1 Announce Type: new Abstract: We study a protocol-level test for weak-label benchmarks: whether benchmark outputs change when the provided evidence is intervened on. Metadata-only sh

Moonwalk: Inverse-Forward Differentiation

Model ReleasesDGX agent

arXiv:2402.14212v4 Announce Type: replace-cross Abstract: Backpropagation's main limitation is its need to store intermediate activations (residuals) during the forward pass, which restricts the depth

Non-normal spectral signatures of instability in neural network training dynamics

Model ReleasesDGX agent

arXiv:2605.23476v1 Announce Type: new Abstract: Training instabilities in deep networks - loss spikes, oscillatory convergence, and gradient pathologies - are empirically prevalent but lack a rigorous

One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents

Model ReleasesDGX agent

arXiv:2605.23652v1 Announce Type: new Abstract: On a 300-persona life-simulation benchmark, pcsp achieves compositional zero-shot persona identification up to 17x above chance, Spearman rho approx 0.7

OnePred: Next-Query Prediction via Recursive Intent Memory in Multi-Turn Conversations

ResearchDGX agent

arXiv:2605.23668v1 Announce Type: cross Abstract: Although large language model (LLM) conversational systems process millions of multi-turn dialogues daily, they remain fundamentally reactive: they re

Online Partitioned Local Depth for semi-supervised applications

Model ReleasesDGX agent

arXiv:2512.15436v2 Announce Type: replace-cross Abstract: We introduce an extension of the partitioned local depth (PaLD) algorithm that is adapted to online applications such as semi-supervised predi

← Previous
1…667668669670671…1042
Next →