AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
91,566Total entries
1Added by human
91,565Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,232 results
Model Releases

Gate AI: LLM Security Benchmark Evaluation Methodology and Results

DGX agent

arXiv:2606.02959v1 Announce Type: new Abstract: Published evaluations of prompt-injection and jailbreak detectors for Large Language Models often suffer from two systematic weaknesses: per-dataset thr

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Geometry-Aware Tabular Diffusion

DGX agent

arXiv:2606.02607v1 Announce Type: cross Abstract: Tabular synthesis is critical for privacy-preserving sharing and augmentation, yet diffusion models rely on implicit mechanisms to capture inter-colum

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Hallucinations as Orthogonal Noise: Inference-Time Manifold Alignment via Dynamic Contextual Orthogonalization

DGX agent

arXiv:2606.03022v1 Announce Type: cross Abstract: Hallucination in Large Language Models (LLMs), characterized by the generation of content inconsistent with contextual facts or logical constraints --

model-releasesarxiv-cs-ai
3 Jun 2026
Applications

Harvey just published a study showing a hybrid setup, open source GLM 5.1 as primary worker, routing to Opus 4.7 only when needed beats pure…

DGX agent

Harvey just published a study showing a hybrid setup, open source GLM 5.1 as primary worker, routing to Opus 4.7 only when needed beats pure Opus 4.7 on quality and costs less. This is the multi-model

applicationsclem-delangue--x
3 Jun 2026
Tutorials

HiSE: A Lightweight Hierarchical Semantic Explainer for Heterogeneous Graph Neural Networks

DGX agent

arXiv:2606.03495v1 Announce Type: new Abstract: Heterogeneous graph neural networks (HGNNs) have demonstrated remarkable performance in modeling complex relational data, however their interpretability

tutorialsarxiv-cs-lg
3 Jun 2026
Model Releases

HyperPatch: Sequential Knowledge Editing Under n-ary Structural Drift

DGX agent

arXiv:2606.03179v1 Announce Type: new Abstract: Large Language Models (LLMs) rely on Knowledge Editing (KE) to maintain temporal validity, yet real-world knowledge is inherently n-ary. We demonstrate

model-releasesarxiv-cs-cl
3 Jun 2026
Local Ai

Learn from Your Mistakes: Tree-like Self-Play for Secure Code LLMs

DGX agent

arXiv:2606.03489v1 Announce Type: cross Abstract: While Large Language Models (LLMs) excel in code generation, they remain prone to replicating subtle yet critical vulnerabilities endemic to their tra

local-aiarxiv-cs-ai
3 Jun 2026
Safety

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs

DGX agent

arXiv:2602.10352v2 Announce Type: replace-cross Abstract: Self-interpretation methods prompt language models to describe their own internal states, but remain unreliable due to hyperparameter sensitiv

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

MemoGen: Can Past Experience Improve Future Text-to-Image Generation?

DGX agent

arXiv:2606.03243v1 Announce Type: new Abstract: Modern text-to-image models have achieved strong visual synthesis, yet remain unreliable when prompts require implicit visual constraints, relational re

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Message Tuning Outshines Graph Prompt Tuning: A Prismatic Space Perspective

DGX agent

arXiv:2606.03290v1 Announce Type: cross Abstract: Graph Foundation Models (GFMs), built upon the Pre-training and Adaptation paradigm, have emerged as a research hotspot in graph learning. For GNN-bas

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

MiMo v2.5 (pro) availability

DGX agent

MiMo-V2.5-Pro is a model available on Hugging Face that was requested to be added to Ollama's cloud models in May 2026. The discussion on r/ollama likely covers the availability status of this Xiaomi-

local-air-ollama
3 Jun 2026
Model Releases

Multi^2: Hierarchical Multi-Agent Decision-Making with LLM-Based Agents in Interactive Environments

DGX agent

arXiv:2606.03698v1 Announce Type: new Abstract: A central goal of large language model (LLM) research is to build agentic systems that can plan, act, and adapt through sustained interaction with dynam

model-releasesarxiv-cs-lg
3 Jun 2026
Local Ai

Nanocoder 1.27.0 - skills, daemon + more 🔥

DGX agent

Nanocoder 1.27.0 is an agentic coding tool available in your terminal that runs on any AI model you choose, whether local models via Ollama or cloud providers like OpenAI and Anthropic. This release i

local-air-ollama
3 Jun 2026
Model Releases

NeuroArmor: Safe-Variant-Guided Representation Consistency for Selective Re-Anchoring in Jailbreak Defense

DGX agent

arXiv:2606.03486v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak attacks that hide harmful intent behind seemingly ordinary requests such as role-play, translatio

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

OpenEAI-Platform: An Open-source Embodied Artificial Intelligence Hardware-Software Unified Platform

DGX agent

arXiv:2606.03392v1 Announce Type: new Abstract: Embodied AI in the real world requires both accurate hardware and robust vision-language-action (VLA) policies. We present OpenEAI-Platform, a fully ope

model-releasesarxiv-cs-ro
3 Jun 2026
Model Releases

Perceive Before Reasoning: A Pre-Reasoning Perception Framework for Efficient and Reliable Proactive Mobile Agents

DGX agent

arXiv:2606.03236v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have substantially advanced mobile agents, yet proactive mobile assistance remains challenging because agents m

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

PrimeSVT: An Automated Memory-aware Pruning Framework with Prioritized Compression Policy for Spiking Vision Transformers

DGX agent

arXiv:2606.03428v1 Announce Type: cross Abstract: The large sizes of Spiking Vision Transformers (SViTs) still hinder their embedded implementation, highlighting the need for model compression. State-

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

Proof-Refactor: Refactoring Generated Formal Proofs into Modular Artifacts

DGX agent

arXiv:2606.03743v1 Announce Type: new Abstract: While Large Language Models (LLMs) have shown strong performance in generating formal proofs, their outputs often remain less readable, modular, maintai

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Qwen-Image-Flash: Beyond Objective Design

DGX agent

arXiv:2606.03746v1 Announce Type: cross Abstract: Few-step distillation has become an effective strategy for accelerating advanced visual generative models, yet prior work has largely focused on disti

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Samudra 2: Scaling Ocean Emulators across Resolutions

DGX agent

arXiv:2606.02610v1 Announce Type: cross Abstract: Ocean general circulation models (OGCMs) are essential to climate science but computationally expensive, limiting ensemble size and forcing scenarios.

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Scalable On-Hardware Training of Quantum Neural Networks and Application to Clinical Data Imputation

DGX agent

arXiv:2606.03517v1 Announce Type: cross Abstract: Training quantum neural networks (QNNs) on quantum hardware is currently bottlenecked by the cost of gradient estimation: standard parameter-shift met

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

SEA-NLI: Natural Language Inference as a Lens into Southeast Asian Cultural Understanding

DGX agent

arXiv:2606.03284v1 Announce Type: new Abstract: Frontier LLMs perform well in Western contexts, but remain poorly tested on underrepresented cultures such as those in Southeast Asia (SEA). Existing NL

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

SeeTraceAct: Visibility-Aware Latent Planning from Cross-Embodiment Demonstration Videos

DGX agent

arXiv:2606.02745v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) are promising general-purpose robot policies, but adapting them to new tasks typically requires costly task-speci

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

SynCred-Bench: Benchmarking Synthetic Credibility in AI-Generated Visual Misinformation

DGX agent

arXiv:2606.03348v1 Announce Type: cross Abstract: Recent generative models can now produce visual artifacts with realistic embedded text and layouts, creating a new misinformation threat: synthetic cr

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Trading Human Curation for Synthetic Augmentation in RLVR

DGX agent

arXiv:2606.03800v1 Announce Type: cross Abstract: The supply of high-quality training tasks is a central bottleneck for reinforcement learning from verifiable rewards (RLVR) on agentic language models

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

TriEval: A Resource-Efficient Pipeline for LLM Bias, Toxicity, and Truthfulness Assessment

DGX agent

arXiv:2606.03036v1 Announce Type: new Abstract: LLMs have evolved from basic chatbots to the backbone of the AI ecosystem, now widely used in healthcare, schools, and government services. The domain-w

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

VidMsg: A Benchmark for Implicit Message Inference in Short Videos

DGX agent

arXiv:2606.03635v1 Announce Type: cross Abstract: Understanding short online videos involves more than identifying visible objects and actions; video makers often include an underlying message or purp

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Weak Diffusion Priors Can Still Achieve Strong Inverse-Problem Performance

DGX agent

arXiv:2601.22443v2 Announce Type: replace-cross Abstract: Can a diffusion model trained on bedrooms recover human faces? Diffusion models are widely used as priors for inverse problems, but standard a

researcharxiv-cs-cv
3 Jun 2026
Tutorials

What Do Students Learn? A Feature-Level Analysis of Dark Knowledge

DGX agent

arXiv:2606.03052v1 Announce Type: new Abstract: Knowledge Distillation (KD) is a powerful tool for model compression, yet the precise mechanisms by which student models acquire feature representations

tutorialsarxiv-cs-lg
3 Jun 2026
Model Releases

Which Defense Closes Which Threat? Attributing OWASP-LLM-Top-10 Coverage and Its Brittleness Under Paraphrasing

DGX agent

arXiv:2606.02822v1 Announce Type: cross Abstract: Production LLM applications stack several defense families -- refusal-phrase filters, token-budget controls, model allowlists, rate limits, tool-regis

model-releasesarxiv-cs-ai
3 Jun 2026
Research

X-RAY: Mapping LLM Reasoning Capability via Formalized and Calibrated Probes

DGX agent

arXiv:2603.05290v2 Announce Type: replace Abstract: Large language models (LLMs) achieve promising performance, yet their ability to reason remains poorly understood. Existing evaluations largely emph

researcharxiv-cs-ai
3 Jun 2026
Model Releases

A Closer Look at In-Distribution vs. Out-of-Distribution Accuracy for Open-Set Test-time Adaptation

DGX agent

arXiv:2606.01973v1 Announce Type: cross Abstract: Open-set test-time adaptation (TTA) updates models on new data in the presence of input shifts and unknown output classes. While recent methods have m

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

A Machine-to-Machine Knowledge-Guided LLM Agent for Generalizable Radiotherapy Treatment Planning

DGX agent

arXiv:2606.00922v1 Announce Type: cross Abstract: In this work, we propose a prototype machine-to-machine (M2M) knowledge-guided Large Language Model (LLM) framework for automated radiotherapy treatme

model-releasesarxiv-cs-ro
2 Jun 2026
Research

A unifying Bayesian framework for adversarial robustness

DGX agent

arXiv:2510.09288v2 Announce Type: replace-cross Abstract: The vulnerability of machine learning models to adversarial attacks remains a critical societal security challenge. Traditional defenses, such

researcharxiv-cs-lg
2 Jun 2026
Model Releases

Anthropic expands Project Glasswing cybersecurity program to 150 more organizations

DGX agent

Anthropic PBC is expanding a program that enables organizations to test their cybersecurity defenses using its Claude Mythos Preview model. The initiative, which is known as Project Glasswing, launche

model-releasessiliconangle
2 Jun 2026
Research

ASKD-Whisper: Adaptive Self-knowledge Distillation for Efficient and Low-Latency Automatic Speech Recognition

DGX agent

arXiv:2601.19919v2 Announce Type: replace-cross Abstract: Knowledge distillation (KD) is one of the most effective paradigms for compressing large-scale foundation models into deployable architectures

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Attention mechanisms and transfer learning for robust peach leaf damage classification under domain shift

DGX agent

arXiv:2606.02045v1 Announce Type: cross Abstract: Artificial intelligence provides a practical framework for crop damage assessment from imagery data, supporting early decision-making in agricultural

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

AXIOM: A Trust-First Neuro-Symbolic Execution Architecture for Verifiable Mathematical Reasoning

DGX agent

arXiv:2606.00671v1 Announce Type: new Abstract: We present AXIOM, a trust-first neuro-symbolic execution architecture for natural-language mathematical reasoning. In AXIOM, the language model function

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Better with Experience: Self-Evolving LLM Agents for Evidence-Grounded Health Community Notes

DGX agent

arXiv:2606.02215v1 Announce Type: new Abstract: Large Language Model (LLM)-augmented Community Notes offer a scalable path for timely, evidence-grounded correction of health misinformation on social p

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Beyond Rigid: Benchmarking Non-Rigid Video Editing

DGX agent

arXiv:2601.18340v2 Announce Type: replace Abstract: As video generation models are increasingly expected to manipulate physical dynamics, there is a growing need to move evaluation beyond appearance f

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

DGX agent

arXiv:2603.19453v2 Announce Type: replace Abstract: We study LLM policy synthesis: using a language model to iteratively generate programmatic agent policies for multi-agent environments. Rather than

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CART: Context-Anchored Recurrent Transformer -- A Parameter-Efficient Architecture with Learned Stability

DGX agent

arXiv:2606.01495v1 Announce Type: cross Abstract: We present CART (Context-Anchored Recurrent Transformer), a parameter-efficient language model that reuses a single shared core block R times across d

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms for LLMs

DGX agent

arXiv:2603.24511v2 Announce Type: replace-cross Abstract: We show that AI agents are capable of discovering novel algorithms for adversarial attacks against LLMs, advancing the state of the art on whi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Codex is becoming a productivity tool for everyone

DGX agent

OpenAI's Codex is evolving beyond code generation to become a general productivity tool accessible to non-programmers for knowledge work tasks. The tool leverages large language models to assist with

model-releasesopenai
2 Jun 2026
Research

Continuous Reasoning for Vision-Language-Action

DGX agent

arXiv:2606.00229v1 Announce Type: cross Abstract: Natural language is a powerful reasoning medium for language and vision-language models, but it is mismatched to the granularity of continuous control

researcharxiv-cs-ai
2 Jun 2026
Model Releases

ContinuousBench: Can Differentially Private Synthetic Text Improve Capabilities?

DGX agent

arXiv:2606.01849v1 Announce Type: cross Abstract: Differentially private (DP) text synthesis promises to unlock sensitive corpora for model training, but it remains unclear whether DP synthetic data t

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CRAM: Centroid-Routing and Adaptive MoE for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2606.02502v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) unify heterogeneous vision-language tasks under a shared generative framework via instruction tuning, yet real-

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CRMA: A Spectrally-Bounded Backbone for Modular Continual Fine-Tuning of LLMs

DGX agent

arXiv:2606.00382v1 Announce Type: new Abstract: Sequential fine-tuning of large language models forces a choice: let the shared substrate keep learning and accept catastrophic forgetting, or freeze it

model-releasesarxiv-cs-lg
2 Jun 2026
← Previous
1…528529530531532…1380
Next →