AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
91,566Total entries
1Added by human
91,565Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,232 results
Model Releases

Cross-Generational Transfer of Adversarial Attacks Reveals Non-Monotonic Safety Alignment in LLMs

DGX agent

arXiv:2606.00813v1 Announce Type: cross Abstract: Safety alignment in LLMs does not improve monotonically across model generations. Studying four generations of Google's Gemma family (7B-31B) with qua

model-releasesarxiv-cs-cl
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CV-Arena: An Open Benchmark for Instructional Computer Vision Problem Solving with Human-AI Collaborative Preferences

DGX agent

arXiv:2606.00931v1 Announce Type: cross Abstract: Instruction-guided image editing is becoming a general interface for visual work, yet existing benchmarks still focus largely on narrow appearance edi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Data Collection for Training Quality-Control AI in Carpet Manufacturing

DGX agent

arXiv:2606.01023v1 Announce Type: cross Abstract: Visual inspection remains the dominant quality-control practice in woven and tufted carpet production, yet it is slow, subjective, and inconsistent at

model-releasesarxiv-cs-ai
2 Jun 2026
Research

DenseMLLM: Standard Multimodal LLMs for Dense Prediction

DGX agent

arXiv:2602.14134v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated exceptional capabilities in high-level visual understanding. However, extending the

researcharxiv-cs-ai
2 Jun 2026
Safety

Dialectics of Alignment: Harnessing Unsafe Knowledge for Dynamic Safety Routing

DGX agent

arXiv:2606.00686v1 Announce Type: new Abstract: The prevailing paradigm in large language model (LLM) alignment operates via erasure, filtering unsafe data or training models to strictly refuse harmfu

safetyarxiv-cs-lg
2 Jun 2026
Safety

DOT-MoE: Differentiable Optimal Transport for MoEfication

DGX agent

arXiv:2606.01666v1 Announce Type: cross Abstract: The scaling of Large Language Models (LLMs) has driven significant performance gains but created substantial challenges in inference efficiency. While

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Dr. DocBench: A Comprehensive Benchmark for Expert-Level and Difficult Document Parsing

DGX agent

arXiv:2606.01393v1 Announce Type: cross Abstract: Document parsing and recognition are fundamental capabilities for vision-language models (VLMs) and document processing systems. However, existing Opt

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Echo State Networks for Time Series Forecasting: Hyperparameter Sweep and Benchmarking

DGX agent

arXiv:2602.03912v4 Announce Type: replace Abstract: This paper investigates the performance of Echo State Networks (ESNs) for univariate forecasting of monthly and quarterly time series from the M4 Fo

model-releasesarxiv-cs-lg
2 Jun 2026
Research

eMoT: evolving Memory-of-Thought via Symbolic Anchoring and Memory Corrosion

DGX agent

arXiv:2606.02054v1 Announce Type: new Abstract: While Large Language Models (LLMs) achieve impressive performance on multi-step reasoning tasks, their reliability is persistently hindered by critical

researcharxiv-cs-ai
2 Jun 2026
Applications

Enhancing BiGRU with a KAN Block for Legal Document Classification and Summarization

DGX agent

arXiv:2606.00116v1 Announce Type: cross Abstract: This study introduces a novel architecture of KAN-based BiGRU model for the task of classification and summarization of legal documents in a low-resou

applicationsarxiv-cs-ai
2 Jun 2026
Model Releases

EuraGovExam: A Multilingual Multimodal Benchmark from Real-World Civil Service Exams

DGX agent

arXiv:2603.27223v2 Announce Type: replace-cross Abstract: We present EuraGovExam, a multilingual and multimodal benchmark sourced from real-world civil service examinations across five representative

model-releasesarxiv-cs-ai
2 Jun 2026
Research

FLaG: Fine-Grained Latent Grouping for Hallucination Detection

DGX agent

arXiv:2606.00301v1 Announce Type: new Abstract: Hallucinations in large language models (LLMs) arise from heterogeneous failure mechanisms, making reliable detection difficult for any single global un

researcharxiv-cs-lg
2 Jun 2026
Model Releases

Flow Matching for Convective-Scale Precipitation Downscaling

DGX agent

arXiv:2606.00281v1 Announce Type: cross Abstract: Generative machine learning is an increasingly important complement to dynamical downscaling for producing high-resolution precipitation projections,

model-releasesarxiv-cs-lg
2 Jun 2026
Local Ai

Flux klein 9b Comic Character Lora?

DGX agent

This post discusses creating or using a LoRA (Low-Rank Adaptation) model compatible with Flux Klein 9b, an AI image generation model, specifically for generating comic book-style characters. The discu

local-air-stablediffusion
2 Jun 2026
Model Releases

From Evaluation to Design: Using Potential Energy Surface Smoothness Metrics to Guide Machine Learning Interatomic Potential Architectures

DGX agent

arXiv:2602.04861v2 Announce Type: replace-cross Abstract: Machine Learning Interatomic Potentials (MLIPs) sometimes fail to reproduce the physical smoothness of the quantum potential energy surface (P

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

From Outliers to Errors: Auditing Pali-to-English LLM Translations with Multi-Reference Adjudication

DGX agent

arXiv:2606.01136v1 Announce Type: new Abstract: Single-score translation metrics can conflate legitimate variation with error, a problem especially acute for classical languages where multiple defensi

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

GABI: Geometry-Aware Boundary Integration for Spacecraft Segmentation

DGX agent

arXiv:2606.00886v1 Announce Type: new Abstract: Accurate segmentation is crucial for autonomous spacecraft, as it directly affects downstream tasks related to 3D situational awareness. The harsh illum

model-releasesarxiv-cs-cv
2 Jun 2026
Research

GateKD: Confidence-Gated Closed-Loop Distillation for Robust Reasoning

DGX agent

arXiv:2605.13136v2 Announce Type: replace Abstract: Distilling multi-step reasoning abilities from large language models (LLMs) into compact student models remains challenging due to noisy rationales,

researcharxiv-cs-cl
2 Jun 2026
Model Releases

Global PIQA: Evaluating Commonsense Reasoning Across 100+ Languages and Cultures

DGX agent

arXiv:2510.24081v2 Announce Type: replace Abstract: To date, there exist almost no culturally-specific evaluation benchmarks for large language models (LLMs) that cover a large number of languages and

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Hierarchical Online Prompt Mutation with Dual-Loop Feedback for Guardrailed Evidence Document Generation: A Production-Evaluation Case Study

DGX agent

arXiv:2606.01472v1 Announce Type: cross Abstract: High-stakes production document-generation systems require language models to be adaptive, evidence-grounded, and auditable. We present HOPM, a hierar

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

HomeFlow: A Data Flywheel for Smart Home Agent Training with Verifiable Simulation

DGX agent

arXiv:2606.01230v1 Announce Type: new Abstract: Large language model agents are moving beyond text-only interaction toward physical-world control, with smart homes as a representative domain. Real dom

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Hybrid Imbalanced Regression Through Unified Data-Level and Algorithm-Level Balancing

DGX agent

arXiv:2606.01221v1 Announce Type: cross Abstract: Imbalanced learning is a critical challenge in machine learning, where underrepresented target values can bias models and degrade prediction performan

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Ideas in Inference-time Scaling can Benefit Generative Pre-training Algorithms

DGX agent

arXiv:2503.07154v3 Announce Type: replace-cross Abstract: Generative pre-training is often framed through a false dichotomy between autoregressive models for discrete signals and diffusion models for

researcharxiv-cs-ai
2 Jun 2026
Model Releases

InsightVQA: High-Dimensional Emotion-Cognitive Visual Question Answering Benchmark

DGX agent

arXiv:2606.02171v1 Announce Type: new Abstract: Visual emotion understanding requires models not only to recognize emotional states, but also to why they arise and perform higher-level cognitive reaso

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

LALE: Lightweight-Transformer Architecture for Land-Cover Estimation

DGX agent

arXiv:2606.02092v1 Announce Type: cross Abstract: Semantic segmentation of remote sensing imagery requires models that capture both global context and local detail under tight computational budgets. P

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

LaSR: Context-Aware Speech Recognition via Latent Reasoning

DGX agent

arXiv:2606.00507v1 Announce Type: new Abstract: Recent advances in Speech Large Language Models (Speech LLMs) have significantly enhanced spoken language understanding and reasoning. However, their co

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

Latent Collaboration in Multi-Agent Systems

DGX agent

arXiv:2511.20639v3 Announce Type: replace-cross Abstract: Multi-agent systems (MAS) extend large language models (LLMs) from independent single-model reasoning to coordinative system-level intelligenc

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

LeAP: Learnable Adaptive Permutation for Feature Selection in Heterogeneous and Sparse Recommender Systems

DGX agent

arXiv:2606.01111v1 Announce Type: new Abstract: Modern industrial recommender systems rely on thousands of heterogeneous features -- ranging from low-dimensional scalars (e.g., statistical value) to h

model-releasesarxiv-cs-lg
2 Jun 2026
Research

Learning from Saturated Data: Signals Beyond Correctness for LLM Training

DGX agent

arXiv:2606.01436v1 Announce Type: new Abstract: The growing capabilities of large language models (LLMs) have led to the saturation of many benchmarks and training datasets used to improve them. Motiv

researcharxiv-cs-cl
2 Jun 2026
Model Releases

LLM Consortium for Software Design Refinement: A Controlled Experiment on Multi-Agent Collaboration Topologies

DGX agent

arXiv:2606.01490v1 Announce Type: cross Abstract: We present a controlled experiment evaluating 12 multi-agent LLM collaboration topologies for software architecture design. Using a 2imes2imes2 factor

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

LLM4Cov: Execution-Aware Agentic Learning for High-coverage Testbench Generation

DGX agent

arXiv:2602.16953v3 Announce Type: replace Abstract: Execution-aware LLM agents offer a promising paradigm for learning from tool feedback, but such feedback can be expensive and slow to obtain, making

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

LLMs for Cardiovascular Risk Prediction from Structured Clinical Data

DGX agent

arXiv:2606.00031v1 Announce Type: cross Abstract: Coronary artery disease (CAD) remains one of the leading causes of death globally, highlighting the need for reliable predictive systems to support ea

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Make a Video Call with LLM: A Measurement Campaign over Six Mainstream Apps

DGX agent

arXiv:2510.00481v2 Announce Type: replace-cross Abstract: In 2025, Large Language Model (LLM) services have launched a new feature -- AI video chat -- allowing users to interact with AI agents via rea

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Med-Scout: Curing MLLMs' Geometric Blindness in Medical Perception via Geometry-Aware RL Post-Training

DGX agent

arXiv:2601.23220v2 Announce Type: replace-cross Abstract: Despite recent Multimodal Large Language Models (MLLMs)' linguistic prowess in medical diagnosis, we find even state-of-the-art MLLMs suffer f

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

@Microsoft @nvidia roundup of links: https://www.latent.space/p/ainews-nvidia-cosmos-3-nemotron-3

DGX agent

This post compiles recent AI news and developments from Microsoft and NVIDIA, likely covering topics such as NVIDIA's Cosmos 3 model and Nemotron 3 framework, along with other significant updates in t

model-releasesswyx--x
2 Jun 2026
Applications

MineDraft: A Framework for Batch Parallel Speculative Decoding

DGX agent

arXiv:2603.18016v2 Announce Type: replace-cross Abstract: Speculative decoding (SD) accelerates large language model inference by using a smaller draft model to propose draft tokens that are subsequen

applicationsarxiv-cs-ai
2 Jun 2026
Model Releases

MixerSENet: A Lightweight Framework for Efficient Hyperspectral Image Classification

DGX agent

arXiv:2606.01700v1 Announce Type: new Abstract: In this paper, a novel framework, MixerSENet, is introduced for hyperspectral image (HSI) classification, designed to address the challenges of computat

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

MMDG-Bench: A Benchmark for Multimodal Domain Generalization

DGX agent

arXiv:2606.00891v1 Announce Type: new Abstract: Multi-modal Domain Generalization (MMDG) seeks to leverage complementary modalities to enhance model robustness on unseen domains. Despite extensive pro

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

MOC: Multi-Order Communication in LLM-based Multi-Agent Systems

DGX agent

arXiv:2606.02359v1 Announce Type: new Abstract: Despite the remarkable progress of Large Language Model (LLM) based Multi-Agent Systems, most research focuses on optimizing coordination topology while

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

MomentKV: Closing the Directional Gap in KV Cache Eviction for Long-Context Inference

DGX agent

arXiv:2606.01563v1 Announce Type: new Abstract: Autoregressive decoding in Transformer-based language models relies on the KV cache, whose memory footprint grows linearly with sequence length and beco

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU Fabrics

DGX agent

arXiv:2606.01502v1 Announce Type: cross Abstract: Frontier LLMs increasingly decide what a query attends to with a sparse-attention indexer that picks a few KV-cache blocks per query: attention's unit

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Multimodal Action Diffusion for Robust End-to-End Autonomous Driving

DGX agent

arXiv:2606.02105v1 Announce Type: new Abstract: End-to-End Autonomous Driving (E2E-AD) systems have largely converged on predicting intermediate trajectory waypoints, delegating final control to hand-

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Multimodal Approaches for Visually-Rich Document Type Classification: A Comparative Analysis

DGX agent

arXiv:2606.02162v1 Announce Type: cross Abstract: Document type classification in visually rich documents remains challenging, as relevant information is distributed across textual, visual, and layout

model-releasesarxiv-cs-ai
2 Jun 2026
Local Ai

Multimodal Function Vectors for Visual Relations

DGX agent

arXiv:2510.02528v2 Announce Type: replace Abstract: Large Multimodal Models (LMMs) demonstrate impressive in-context learning abilities from few multimodal demonstrations, yet the internal mechanisms

local-aiarxiv-cs-ai
2 Jun 2026
Research

Not All Explanations Simulate Equally: Comparing Verbalized Feature Attributions and Self-Generated Rationales

DGX agent

arXiv:2606.01148v1 Announce Type: new Abstract: Natural-language explanations are often treated as a unified interface for understanding model behavior, but different explanation sources may support s

researcharxiv-cs-cl
2 Jun 2026
Model Releases

OneVLA: A Unified Framework for Embodied Tasks

DGX agent

arXiv:2606.01241v1 Announce Type: new Abstract: Navigation and manipulation are fundamental capabilities of embodied intelligence, enabling robots to interpret natural language commands and interact p

model-releasesarxiv-cs-ro
2 Jun 2026
Safety

OPD+: Rethinking the Advantage Design for On-Policy Distillation

DGX agent

arXiv:2606.01039v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a widely used technique to transfer capabilities from capable teacher language models to the base student models, and

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Parameter-efficient Dual-encoder Architecture with Differentiable Choquet Integral Fusion for Underwater Acoustic Classification

DGX agent

arXiv:2606.02341v1 Announce Type: cross Abstract: Underwater acoustic classification has a wide array of oceanic applications, but faces challenges due to an increasingly complex acoustic environment.

model-releasesarxiv-cs-lg
2 Jun 2026
← Previous
1…529530531532533…1380
Next →