AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
4 Jun 2026

CodegenBench: Can LLMs Write Efficient Code Across Architectures?

Model ReleasesDGX agent

arXiv:2606.04023v1 Announce Type: cross Abstract: While large language models (LLMs) have been extensively evaluated on code generation tasks for general-purpose programming and GPU-accelerated enviro

COMBINER: Composed Image Retrieval Guided by Attribute-based Neighbor Relations

Model ReleasesDGX agent

arXiv:2606.04604v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) represents a challenging retrieval task that targets locating specific images through multimodal inputs. Despite recent p

Do Transformers Need Three Projections? Systematic Study of QKV Variants

Model ReleasesDGX agent

arXiv:2606.04032v1 Announce Type: cross Abstract: Transformers have become the standard solution for various AI tasks, with the query, key, and value (QKV) attention formulation playing a central role

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Dreaming: Better memory for a more helpful ChatGPT

Model ReleasesDGX agent

OpenAI's 'Dreaming' feature enhances ChatGPT's memory capabilities by allowing the model to retain and leverage information from previous conversations to provide more personalized and contextually aw

FLAGG: Flexible Autoregressive Graph Generation

SafetyDGX agent

arXiv:2606.05067v1 Announce Type: new Abstract: The Deep Graph Generation's panorama spans two extremes: one-shot and sequential models. The former generates nodes and edges jointly, while the latter

GeM-NR: Geometry-Aware Multi-View Editing for Nonrigid Scene Changes

Model ReleasesDGX agent

arXiv:2606.05142v1 Announce Type: cross Abstract: Recent developments in multi-view image editing with generative models have brought us a step closer toward general 3D content generation and customiz

How dynamic workflows allow Claude Code to handle whole new types of tasks https://x.com/trq212/status/2061907337154367865

Model ReleasesDGX agent

Dynamic workflows in Claude Code enable the model to handle complex, multi-step tasks by allowing execution flows to adapt based on intermediate results rather than following fixed paths. This capabil

Iliad (Troy) trailer made by Grok Imagine 1.5, which was just released

Model ReleasesDGX agent

Elon Musk shared a trailer for 'Iliad (Troy)' created using Grok Imagine 1.5, Xai's newly released text-to-image generation model. The post demonstrates the capabilities of the latest version of Grok'

Impostor: An Agent-Curated Benchmark for Realistic AIGC Manipulation Localization

Model ReleasesDGX agent

arXiv:2606.04545v1 Announce Type: new Abstract: Recent advances in generative image editing have improved the realism and controllability of localized image manipulation, raising new challenges for im

Inclusion-of-Thoughts: Mitigating Preference Instability via Purifying the Decision Space

ResearchDGX agent

arXiv:2604.04944v2 Announce Type: replace-cross Abstract: Multiple-choice questions (MCQs) are widely used to evaluate large language models (LLMs). However, LLMs remain vulnerable to the presence of

InstantRetouch: Efficient and High-Fidelity Instruction-Guided Image Retouching with Bilateral Space

Model ReleasesDGX agent

arXiv:2606.05071v1 Announce Type: new Abstract: Language-guided photo retouching aims to adjust color and tone while preserving geometry and texture. Recently, diffusion-based retouching shows a super

Invariant Gradient Alignment for Robust Reasoning Distillation

Model ReleasesDGX agent

arXiv:2606.05025v1 Announce Type: cross Abstract: Large language models (LLMs) suffer from shortcut learning: they systematically fail on out-of-distribution (OOD) inputs whose semantic surface differ

L^3: Large Lookup Layers

ResearchDGX agent

arXiv:2601.21461v3 Announce Type: replace-cross Abstract: Modern sparse language models typically achieve sparsity through Mixture-of-Experts (MoE) layers, which dynamically route tokens to dense MLP

MeshTok: Efficient Multi-Scale Tokenization for Scalable PDE Transformers

Model ReleasesDGX agent

arXiv:2606.04366v1 Announce Type: new Abstract: Conventional patchified Transformers operate on uniform spatial partitions, distributing computational effort evenly across the domain irrespective of l

MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation

Local AiDGX agent

arXiv:2606.04688v1 Announce Type: new Abstract: Autoregressive mesh generation has gained attention by tokenizing meshes into sequences and training models in a language-modeling fashion. However, exi

MetaPoint: Unlocking Precise Spatial Control in Agentic Visual Generation

AgentsDGX agent

arXiv:2606.05031v1 Announce Type: new Abstract: Generative visual models fundamentally struggle with precise spatial control. This arises from a core disconnect: models can process textual description

MusaCoder: Native GPU Kernel Generation with Full-Stack Training on Moore Threads GPU

SafetyDGX agent

arXiv:2606.04847v1 Announce Type: cross Abstract: Native GPU kernel generation turns high-level tensor programs into executable, efficient low-level code. Existing Large Language Models (LLMs) struggl

PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?

Model ReleasesDGX agent

arXiv:2602.01146v2 Announce Type: replace Abstract: Conversational assistants are increasingly integrating long-term memory with large language models (LLMs). This persistence of memories, e.g., the u

RAVQ-HoloNet: Rate-Adaptive Vector-Quantized Hologram Compression

Model ReleasesDGX agent

arXiv:2511.21035v2 Announce Type: replace Abstract: Holography offers significant potential for AR/VR applications. However, its adoption is limited by the high demand for data compression. Existing d

Scaling AI Agents: A Step-by-Step Guide to Deploying ADK on GKE Autopilot

Model ReleasesDGX agent

While building AI agents locally using Google’s Agent Development Kit (ADK) is an excellent way to prototype, production-ready agents require a robust, scalable infrastructure. For developers looking

Shifting the Breaking Point of Flow Matching for Multi-Instance Editing

Model ReleasesDGX agent

arXiv:2602.08749v3 Announce Type: replace Abstract: Flow matching models have recently emerged as an efficient alternative to diffusion, especially for text-guided image generation and editing, offeri

Signed Dual Attention: Capturing Signed Dependencies in Time Series Forecasting

Model ReleasesDGX agent

arXiv:2606.04833v1 Announce Type: cross Abstract: Initially developed for natural language processing, Transformer architectures and attention mechanisms are now central to a wide range of deep learni

SMADE-IE: Sparse Multi-Agent Framework with Evidence-Driven Debate for Zero-Shot Information Extraction

Model ReleasesDGX agent

arXiv:2606.04691v1 Announce Type: new Abstract: Zero-shot information extraction (IE) with large language models (LLMs) has attracted increasing attention due to its flexibility in adapting to new sch

Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2407.03956v3 Announce Type: replace-cross Abstract: Prior research has enhanced the ability of Large Language Models (LLMs) to solve logic puzzles using techniques such as chain-of-thought promp

StandardE2E: A Unified Framework for End-to-End Autonomous Driving Datasets

Model ReleasesDGX agent

arXiv:2606.04271v1 Announce Type: cross Abstract: Autonomous driving has shifted from modular perception-prediction-planning stacks toward end-to-end (E2E) models that map sensor inputs directly to ve

Structure-Aware Prediction of PROTAC-Mediated Protein Degradability via Graph Neural Networks

Model ReleasesDGX agent

arXiv:2606.04021v1 Announce Type: cross Abstract: Proteolysis-targeting chimeras (PROTACs) can selectively degrade disease-causing proteins, yet predicting which targets are amenable to degradation re

TaDA: Calibrated Probe Gating for Task-Domain LoRA Merging

Model ReleasesDGX agent

arXiv:2606.05016v1 Announce Type: new Abstract: Combining a task LoRA adapter with a domain LoRA adapter into a single unified model is a practical yet largely unexplored challenge. Existing methods t

TANDEM: Bi-Level Data Mixture Optimization with Twin Networks

ResearchDGX agent

arXiv:2606.04401v1 Announce Type: new Abstract: The capabilities of large language models (LLMs) significantly depend on training data drawn from various domains. Optimizing domain-specific mixture ra

Thinking Through Signs: PEEL as a Semiotic Scaffolding for Epistemically Accountable AI-Enabled Research

Model ReleasesDGX agent

arXiv:2606.04152v1 Announce Type: new Abstract: Large language models are reshaping research practice while quietly eroding researchers epistemic accountability. This commentary introduces PEEL - Prot

Today I'm launching a new project called SynthTraces 🔥 It is a minimal codebase to generate synthetic coding agent session traces using Pi …

Model ReleasesDGX agent

Today I'm launching a new project called SynthTraces 🔥 It is a minimal codebase to generate synthetic coding agent session traces using Pi (from @badlogicgames) I wanted a large number of coding-agent

Toward Autonomous O-RAN: A Multi-Scale Agentic AI Framework for Real-Time Network Control and Management

AgentsDGX agent

arXiv:2602.14117v2 Announce Type: replace-cross Abstract: Open Radio Access Networks (O-RAN) promise flexible 6G network access through disaggregated, software-driven components and open interfaces, b

Toward Pre-Deployment Assurance for Enterprise AI Agents: Ontology-Grounded Simulation and Trust Certification

Model ReleasesDGX agent

arXiv:2606.04037v1 Announce Type: new Abstract: Pre-deployment verification of enterprise artificial intelligence (AI) agents remains a critical gap between large language model (LLM) capability bench

Towards Efficient and Evidence-grounded Mobility Prediction with LLM-Driven Agent

Model ReleasesDGX agent

arXiv:2606.05130v1 Announce Type: cross Abstract: Individual-level mobility prediction is central to urban simulation, transportation planning, and policy analysis. Supervised sequence models achieve

Trivium: Temporal Regret as a First-Class Objective for Causal-Memory Controllers

AgentsDGX agent

arXiv:2606.04421v1 Announce Type: new Abstract: Many current agentic systems and LLM pipelines correct mistakes by optimizing outcome reward. This addresses only the what of failure: when an outcome d

Unpredictable Safety: Domain-Dependent Compliance and the Transparency Gap in Open-Weight LLMs

Model ReleasesDGX agent

arXiv:2606.04035v1 Announce Type: cross Abstract: We present a systematic study of domain-dependent safety behavior in open-weight LLMs: 7 standardized experiments across 7 ethical domains, testing 5

We're presenting ParseBench at CVPR 2026! ParseBench is the most comprehensive document understanding benchmark for VLMs. ✅ It contains 2k p…

Model ReleasesDGX agent

We're presenting ParseBench at CVPR 2026! ParseBench is the most comprehensive document understanding benchmark for VLMs. ✅ It contains 2k pages of real-world enterprise documents ✅ It has comprehensi

What If Prompt Injection Never Left? Exploring Cross-Session Stored Prompt Injection in Agentic Systems

Model ReleasesDGX agent

arXiv:2606.04425v1 Announce Type: cross Abstract: Modern agentic systems transform LLMs from session-bounded assistants into stateful systems that persist and evolve shared world state across sessions

When Seeing Is Not Believing -- A Benchmark for Search-Grounded Video Misinformation Detection

Model ReleasesDGX agent

arXiv:2606.04098v1 Announce Type: new Abstract: Video misinformation increasingly operates at the semantic and evidential level: authentic footage may be selectively edited, temporally reordered, spli

When you burn so much money you run out of options…

Model ReleasesDGX agent

When you burn so much money you run out of options… Anthropic co-founder and President Daniela Amodei said the high cost of developing AI models is driving firms like hers to look to the public market

Where's gemma4:12b?

Local AiDGX agent

Gemma 4 12B is the first medium-sized, encoder-free multimodal model capable of natively ingesting audio and video , recently released by Google. The model is available on Ollama with 11.7M downloads

3 Jun 2026

Analytical Evaluation of DCA Convergence Properties for Minimizing Prediction Functions of Gaussian RBF Support Vector Regression

Model ReleasesDGX agent

arXiv:2606.03559v1 Announce Type: new Abstract: For nonconvex optimization problems whose objective is the prediction function of a trained Support Vector Regression (SVR) model with the Gaussian radi

AnyAudio-Judge: A Dynamic Rubric-Based Benchmark and Evaluator for Audio Instruction Following

Model ReleasesDGX agent

arXiv:2606.03116v1 Announce Type: cross Abstract: The rapid advancement of instruction-guided audio generation has highlighted the critical need for robust alignment evaluation. Current automated eval

Attention Calibration for Position-Fair Dense Information Retrieval

SafetyDGX agent

arXiv:2606.02737v1 Announce Type: cross Abstract: Dense retrieval models exhibit positional bias: retrieval effectiveness degrades when relevant information appears later in a passage (Zeng et al., 20

Beyond Ideal Instruction: A Comprehensive Framework for Evaluating LLMs in Realistic Interactions

Model ReleasesDGX agent

arXiv:2606.03318v1 Announce Type: new Abstract: Despite great advances in tool-use capabilities of large language models (LLMs), existing evaluation benchmarks struggle to fully align with real-world

Calibration Data Trade-offs Across Capability Dimensions: Why Multi-Source Mixing Matters for High-Sparsity LLM Pruning

Model ReleasesDGX agent

arXiv:2606.03328v1 Announce Type: cross Abstract: Post-training pruning compresses large language models to high sparsity using a small unlabelled calibration set, and recent work has concluded that t

Consistency Training Can Entrench Misalignment

SafetyDGX agent

arXiv:2606.03810v1 Announce Type: cross Abstract: Consistency training encourages a model to produce similar outputs across related inputs or sampling procedures. Such methods are simple, scalable, an

Demystifying Pipeline Parallelism: First Theory for PipeDream

Local AiDGX agent

arXiv:2606.03498v1 Announce Type: new Abstract: Training modern machine learning models increasingly requires computation to be distributed across many accelerators. Data parallelism remains the defau

Diagnosis of Human Object Interaction Detectors for Real World Educational Applications

Model ReleasesDGX agent

arXiv:2606.02789v1 Announce Type: new Abstract: Human-object interaction (HOI) recognition is critical for automatically analyzing student behavior in complex educational environments. Although state-

Evaluating and Calibrating LLM Confidence on Questions with Multiple Correct Answers

Model ReleasesDGX agent

arXiv:2602.07842v2 Announce Type: replace Abstract: Confidence calibration is essential for making large language models (LLMs) reliable, yet existing training-free methods have been primarily studied

Fast and Expressive Multi-Byte Prediction with Probabilistic Circuits

ResearchDGX agent

arXiv:2511.11346v2 Announce Type: replace Abstract: Multi-token prediction (MTP) is a prominent strategy to significantly speed up generation in large language models (LLMs), especially in byte-level

Fast Organic Crystal Structure Prediction with Unit Cell Flow Matching

SafetyDGX agent

arXiv:2606.03199v1 Announce Type: new Abstract: Organic crystal structure prediction (CSP) is a requirement for computational modelling of organic solids, but traditionally costs several CPU-years per

FlashMLA-ETAP: Efficient Transpose Attention Pipeline for Accelerating MLA Inference on NVIDIA H20 GPUs

Model ReleasesDGX agent

arXiv:2506.01969v3 Announce Type: replace-cross Abstract: Efficient inference of Multi-Head Latent Attention (MLA) is challenged by deploying the DeepSeek-R1 671B model on a single Multi-GPU server. T

From Long News to Accurate Forecast: Importance-Aware Fusion and PRM-Guided Reflection for Time Series Forecasting

Model ReleasesDGX agent

arXiv:2606.03097v1 Announce Type: new Abstract: Incorporating news into time series forecasting is appealing because news can reveal abrupt exogenous events that historical values alone cannot recover

Gate AI: LLM Security Benchmark Evaluation Methodology and Results

Model ReleasesDGX agent

arXiv:2606.02959v1 Announce Type: new Abstract: Published evaluations of prompt-injection and jailbreak detectors for Large Language Models often suffer from two systematic weaknesses: per-dataset thr

Geometry-Aware Tabular Diffusion

Model ReleasesDGX agent

arXiv:2606.02607v1 Announce Type: cross Abstract: Tabular synthesis is critical for privacy-preserving sharing and augmentation, yet diffusion models rely on implicit mechanisms to capture inter-colum

Hallucinations as Orthogonal Noise: Inference-Time Manifold Alignment via Dynamic Contextual Orthogonalization

Model ReleasesDGX agent

arXiv:2606.03022v1 Announce Type: cross Abstract: Hallucination in Large Language Models (LLMs), characterized by the generation of content inconsistent with contextual facts or logical constraints --

Harvey just published a study showing a hybrid setup, open source GLM 5.1 as primary worker, routing to Opus 4.7 only when needed beats pure…

ApplicationsDGX agent

Harvey just published a study showing a hybrid setup, open source GLM 5.1 as primary worker, routing to Opus 4.7 only when needed beats pure Opus 4.7 on quality and costs less. This is the multi-model

HiSE: A Lightweight Hierarchical Semantic Explainer for Heterogeneous Graph Neural Networks

TutorialsDGX agent

arXiv:2606.03495v1 Announce Type: new Abstract: Heterogeneous graph neural networks (HGNNs) have demonstrated remarkable performance in modeling complex relational data, however their interpretability

HyperPatch: Sequential Knowledge Editing Under n-ary Structural Drift

Model ReleasesDGX agent

arXiv:2606.03179v1 Announce Type: new Abstract: Large Language Models (LLMs) rely on Knowledge Editing (KE) to maintain temporal validity, yet real-world knowledge is inherently n-ary. We demonstrate

Learn from Your Mistakes: Tree-like Self-Play for Secure Code LLMs

Local AiDGX agent

arXiv:2606.03489v1 Announce Type: cross Abstract: While Large Language Models (LLMs) excel in code generation, they remain prone to replicating subtle yet critical vulnerabilities endemic to their tra

← Previous
1…399400401402403…1053
Next →