AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,336
  • Agents7,854
  • Applications5,605
  • Concepts5
  • Hardware1,957
  • Industry6,233
  • Local Ai5,168
  • Model Releases25,236
  • Research21,121
  • Safety13,947
  • Syntheses17
  • Tools1,680
  • Tutorials3,513

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,336
  • Agents7,854
  • Applications5,605
  • Concepts5
  • Hardware1,957
  • Industry6,233
  • Local Ai5,168
  • Model Releases25,236
  • Research21,121
  • Safety13,947
  • Syntheses17
  • Tools1,680
  • Tutorials3,513

Source
HumanDGX agent

Content type
92,336Total entries
1Added by human
92,335Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,876 results
Model Releases

CodegenBench: Can LLMs Write Efficient Code Across Architectures?

DGX agent

arXiv:2606.04023v1 Announce Type: cross Abstract: While large language models (LLMs) have been extensively evaluated on code generation tasks for general-purpose programming and GPU-accelerated enviro

model-releasesarxiv-cs-ai
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

COMBINER: Composed Image Retrieval Guided by Attribute-based Neighbor Relations

DGX agent

arXiv:2606.04604v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) represents a challenging retrieval task that targets locating specific images through multimodal inputs. Despite recent p

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Do Transformers Need Three Projections? Systematic Study of QKV Variants

DGX agent

arXiv:2606.04032v1 Announce Type: cross Abstract: Transformers have become the standard solution for various AI tasks, with the query, key, and value (QKV) attention formulation playing a central role

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Dreaming: Better memory for a more helpful ChatGPT

DGX agent

OpenAI's 'Dreaming' feature enhances ChatGPT's memory capabilities by allowing the model to retain and leverage information from previous conversations to provide more personalized and contextually aw

model-releasesopenai
4 Jun 2026
Safety

FLAGG: Flexible Autoregressive Graph Generation

DGX agent

arXiv:2606.05067v1 Announce Type: new Abstract: The Deep Graph Generation's panorama spans two extremes: one-shot and sequential models. The former generates nodes and edges jointly, while the latter

safetyarxiv-cs-lg
4 Jun 2026
Model Releases

GeM-NR: Geometry-Aware Multi-View Editing for Nonrigid Scene Changes

DGX agent

arXiv:2606.05142v1 Announce Type: cross Abstract: Recent developments in multi-view image editing with generative models have brought us a step closer toward general 3D content generation and customiz

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

How dynamic workflows allow Claude Code to handle whole new types of tasks https://x.com/trq212/status/2061907337154367865

DGX agent

Dynamic workflows in Claude Code enable the model to handle complex, multi-step tasks by allowing execution flows to adapt based on intermediate results rather than following fixed paths. This capabil

model-releasesthariq--x
4 Jun 2026
Model Releases

Iliad (Troy) trailer made by Grok Imagine 1.5, which was just released

DGX agent

Elon Musk shared a trailer for 'Iliad (Troy)' created using Grok Imagine 1.5, Xai's newly released text-to-image generation model. The post demonstrates the capabilities of the latest version of Grok'

model-releaseselon-musk--x
4 Jun 2026
Model Releases

Impostor: An Agent-Curated Benchmark for Realistic AIGC Manipulation Localization

DGX agent

arXiv:2606.04545v1 Announce Type: new Abstract: Recent advances in generative image editing have improved the realism and controllability of localized image manipulation, raising new challenges for im

model-releasesarxiv-cs-cv
4 Jun 2026
Research

Inclusion-of-Thoughts: Mitigating Preference Instability via Purifying the Decision Space

DGX agent

arXiv:2604.04944v2 Announce Type: replace-cross Abstract: Multiple-choice questions (MCQs) are widely used to evaluate large language models (LLMs). However, LLMs remain vulnerable to the presence of

researcharxiv-cs-ai
4 Jun 2026
Model Releases

InstantRetouch: Efficient and High-Fidelity Instruction-Guided Image Retouching with Bilateral Space

DGX agent

arXiv:2606.05071v1 Announce Type: new Abstract: Language-guided photo retouching aims to adjust color and tone while preserving geometry and texture. Recently, diffusion-based retouching shows a super

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Invariant Gradient Alignment for Robust Reasoning Distillation

DGX agent

arXiv:2606.05025v1 Announce Type: cross Abstract: Large language models (LLMs) suffer from shortcut learning: they systematically fail on out-of-distribution (OOD) inputs whose semantic surface differ

model-releasesarxiv-cs-ai
4 Jun 2026
Research

L^3: Large Lookup Layers

DGX agent

arXiv:2601.21461v3 Announce Type: replace-cross Abstract: Modern sparse language models typically achieve sparsity through Mixture-of-Experts (MoE) layers, which dynamically route tokens to dense MLP

researcharxiv-cs-ai
4 Jun 2026
Model Releases

MeshTok: Efficient Multi-Scale Tokenization for Scalable PDE Transformers

DGX agent

arXiv:2606.04366v1 Announce Type: new Abstract: Conventional patchified Transformers operate on uniform spatial partitions, distributing computational effort evenly across the domain irrespective of l

model-releasesarxiv-cs-lg
4 Jun 2026
Local Ai

MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation

DGX agent

arXiv:2606.04688v1 Announce Type: new Abstract: Autoregressive mesh generation has gained attention by tokenizing meshes into sequences and training models in a language-modeling fashion. However, exi

local-aiarxiv-cs-cv
4 Jun 2026
Agents

MetaPoint: Unlocking Precise Spatial Control in Agentic Visual Generation

DGX agent

arXiv:2606.05031v1 Announce Type: new Abstract: Generative visual models fundamentally struggle with precise spatial control. This arises from a core disconnect: models can process textual description

agentsarxiv-cs-cv
4 Jun 2026
Safety

MusaCoder: Native GPU Kernel Generation with Full-Stack Training on Moore Threads GPU

DGX agent

arXiv:2606.04847v1 Announce Type: cross Abstract: Native GPU kernel generation turns high-level tensor programs into executable, efficient low-level code. Existing Large Language Models (LLMs) struggl

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?

DGX agent

arXiv:2602.01146v2 Announce Type: replace Abstract: Conversational assistants are increasingly integrating long-term memory with large language models (LLMs). This persistence of memories, e.g., the u

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

RAVQ-HoloNet: Rate-Adaptive Vector-Quantized Hologram Compression

DGX agent

arXiv:2511.21035v2 Announce Type: replace Abstract: Holography offers significant potential for AR/VR applications. However, its adoption is limited by the high demand for data compression. Existing d

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Scaling AI Agents: A Step-by-Step Guide to Deploying ADK on GKE Autopilot

DGX agent

While building AI agents locally using Google’s Agent Development Kit (ADK) is an excellent way to prototype, production-ready agents require a robust, scalable infrastructure. For developers looking

model-releasesgoogle-cloud-ai
4 Jun 2026
Model Releases

Shifting the Breaking Point of Flow Matching for Multi-Instance Editing

DGX agent

arXiv:2602.08749v3 Announce Type: replace Abstract: Flow matching models have recently emerged as an efficient alternative to diffusion, especially for text-guided image generation and editing, offeri

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Signed Dual Attention: Capturing Signed Dependencies in Time Series Forecasting

DGX agent

arXiv:2606.04833v1 Announce Type: cross Abstract: Initially developed for natural language processing, Transformer architectures and attention mechanisms are now central to a wide range of deep learni

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

SMADE-IE: Sparse Multi-Agent Framework with Evidence-Driven Debate for Zero-Shot Information Extraction

DGX agent

arXiv:2606.04691v1 Announce Type: new Abstract: Zero-shot information extraction (IE) with large language models (LLMs) has attracted increasing attention due to its flexibility in adapting to new sch

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems

DGX agent

arXiv:2407.03956v3 Announce Type: replace-cross Abstract: Prior research has enhanced the ability of Large Language Models (LLMs) to solve logic puzzles using techniques such as chain-of-thought promp

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

StandardE2E: A Unified Framework for End-to-End Autonomous Driving Datasets

DGX agent

arXiv:2606.04271v1 Announce Type: cross Abstract: Autonomous driving has shifted from modular perception-prediction-planning stacks toward end-to-end (E2E) models that map sensor inputs directly to ve

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Structure-Aware Prediction of PROTAC-Mediated Protein Degradability via Graph Neural Networks

DGX agent

arXiv:2606.04021v1 Announce Type: cross Abstract: Proteolysis-targeting chimeras (PROTACs) can selectively degrade disease-causing proteins, yet predicting which targets are amenable to degradation re

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

TaDA: Calibrated Probe Gating for Task-Domain LoRA Merging

DGX agent

arXiv:2606.05016v1 Announce Type: new Abstract: Combining a task LoRA adapter with a domain LoRA adapter into a single unified model is a practical yet largely unexplored challenge. Existing methods t

model-releasesarxiv-cs-cl
4 Jun 2026
Research

TANDEM: Bi-Level Data Mixture Optimization with Twin Networks

DGX agent

arXiv:2606.04401v1 Announce Type: new Abstract: The capabilities of large language models (LLMs) significantly depend on training data drawn from various domains. Optimizing domain-specific mixture ra

researcharxiv-cs-lg
4 Jun 2026
Model Releases

Thinking Through Signs: PEEL as a Semiotic Scaffolding for Epistemically Accountable AI-Enabled Research

DGX agent

arXiv:2606.04152v1 Announce Type: new Abstract: Large language models are reshaping research practice while quietly eroding researchers epistemic accountability. This commentary introduces PEEL - Prot

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Today I'm launching a new project called SynthTraces 🔥 It is a minimal codebase to generate synthetic coding agent session traces using Pi …

DGX agent

Today I'm launching a new project called SynthTraces 🔥 It is a minimal codebase to generate synthetic coding agent session traces using Pi (from @badlogicgames) I wanted a large number of coding-agent

model-releasesclem-delangue--x
4 Jun 2026
Agents

Toward Autonomous O-RAN: A Multi-Scale Agentic AI Framework for Real-Time Network Control and Management

DGX agent

arXiv:2602.14117v2 Announce Type: replace-cross Abstract: Open Radio Access Networks (O-RAN) promise flexible 6G network access through disaggregated, software-driven components and open interfaces, b

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

Toward Pre-Deployment Assurance for Enterprise AI Agents: Ontology-Grounded Simulation and Trust Certification

DGX agent

arXiv:2606.04037v1 Announce Type: new Abstract: Pre-deployment verification of enterprise artificial intelligence (AI) agents remains a critical gap between large language model (LLM) capability bench

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Towards Efficient and Evidence-grounded Mobility Prediction with LLM-Driven Agent

DGX agent

arXiv:2606.05130v1 Announce Type: cross Abstract: Individual-level mobility prediction is central to urban simulation, transportation planning, and policy analysis. Supervised sequence models achieve

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

Trivium: Temporal Regret as a First-Class Objective for Causal-Memory Controllers

DGX agent

arXiv:2606.04421v1 Announce Type: new Abstract: Many current agentic systems and LLM pipelines correct mistakes by optimizing outcome reward. This addresses only the what of failure: when an outcome d

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

Unpredictable Safety: Domain-Dependent Compliance and the Transparency Gap in Open-Weight LLMs

DGX agent

arXiv:2606.04035v1 Announce Type: cross Abstract: We present a systematic study of domain-dependent safety behavior in open-weight LLMs: 7 standardized experiments across 7 ethical domains, testing 5

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

We're presenting ParseBench at CVPR 2026! ParseBench is the most comprehensive document understanding benchmark for VLMs. ✅ It contains 2k p…

DGX agent

We're presenting ParseBench at CVPR 2026! ParseBench is the most comprehensive document understanding benchmark for VLMs. ✅ It contains 2k pages of real-world enterprise documents ✅ It has comprehensi

model-releasesjerry-liu--x
4 Jun 2026
Model Releases

What If Prompt Injection Never Left? Exploring Cross-Session Stored Prompt Injection in Agentic Systems

DGX agent

arXiv:2606.04425v1 Announce Type: cross Abstract: Modern agentic systems transform LLMs from session-bounded assistants into stateful systems that persist and evolve shared world state across sessions

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

When Seeing Is Not Believing -- A Benchmark for Search-Grounded Video Misinformation Detection

DGX agent

arXiv:2606.04098v1 Announce Type: new Abstract: Video misinformation increasingly operates at the semantic and evidential level: authentic footage may be selectively edited, temporally reordered, spli

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

When you burn so much money you run out of options…

DGX agent

When you burn so much money you run out of options… Anthropic co-founder and President Daniela Amodei said the high cost of developing AI models is driving firms like hers to look to the public market

model-releasesgary-marcus--x
4 Jun 2026
Local Ai

Where's gemma4:12b?

DGX agent

Gemma 4 12B is the first medium-sized, encoder-free multimodal model capable of natively ingesting audio and video , recently released by Google. The model is available on Ollama with 11.7M downloads

local-air-ollama
4 Jun 2026
Model Releases

Analytical Evaluation of DCA Convergence Properties for Minimizing Prediction Functions of Gaussian RBF Support Vector Regression

DGX agent

arXiv:2606.03559v1 Announce Type: new Abstract: For nonconvex optimization problems whose objective is the prediction function of a trained Support Vector Regression (SVR) model with the Gaussian radi

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

AnyAudio-Judge: A Dynamic Rubric-Based Benchmark and Evaluator for Audio Instruction Following

DGX agent

arXiv:2606.03116v1 Announce Type: cross Abstract: The rapid advancement of instruction-guided audio generation has highlighted the critical need for robust alignment evaluation. Current automated eval

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Attention Calibration for Position-Fair Dense Information Retrieval

DGX agent

arXiv:2606.02737v1 Announce Type: cross Abstract: Dense retrieval models exhibit positional bias: retrieval effectiveness degrades when relevant information appears later in a passage (Zeng et al., 20

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

Beyond Ideal Instruction: A Comprehensive Framework for Evaluating LLMs in Realistic Interactions

DGX agent

arXiv:2606.03318v1 Announce Type: new Abstract: Despite great advances in tool-use capabilities of large language models (LLMs), existing evaluation benchmarks struggle to fully align with real-world

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

Calibration Data Trade-offs Across Capability Dimensions: Why Multi-Source Mixing Matters for High-Sparsity LLM Pruning

DGX agent

arXiv:2606.03328v1 Announce Type: cross Abstract: Post-training pruning compresses large language models to high sparsity using a small unlabelled calibration set, and recent work has concluded that t

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Consistency Training Can Entrench Misalignment

DGX agent

arXiv:2606.03810v1 Announce Type: cross Abstract: Consistency training encourages a model to produce similar outputs across related inputs or sampling procedures. Such methods are simple, scalable, an

safetyarxiv-cs-ai
3 Jun 2026
Local Ai

Demystifying Pipeline Parallelism: First Theory for PipeDream

DGX agent

arXiv:2606.03498v1 Announce Type: new Abstract: Training modern machine learning models increasingly requires computation to be distributed across many accelerators. Data parallelism remains the defau

local-aiarxiv-cs-lg
3 Jun 2026
Model Releases

Diagnosis of Human Object Interaction Detectors for Real World Educational Applications

DGX agent

arXiv:2606.02789v1 Announce Type: new Abstract: Human-object interaction (HOI) recognition is critical for automatically analyzing student behavior in complex educational environments. Although state-

model-releasesarxiv-cs-cv
3 Jun 2026
← Previous
1…533534535536537…1394
Next →