AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,191 results
Model Releases

SigmaScale: LLM Compression with SVD-based Low-Rank Decomposition and Learned Scaling Matrices

DGX agent

arXiv:2606.07098v1 Announce Type: new Abstract: We present SigmaScale, a method for learning auxiliary scaling matrices S to aid truncated Singular Value Decomposition (SVD) based Large Language Model

model-releasesarxiv-cs-cl
8 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

STREAM: Stochastic Riemannian Flow Matching with Anisotropic Decoder for Digital Histopathology Image Generation

DGX agent

arXiv:2606.07036v1 Announce Type: cross Abstract: Synthetic histopathology image generation addresses critical challenges in computational pathology, including patient privacy and the growing need for

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

TEVI: Text-Conditioned Editing of Visual Representations via Sparse Autoencoders for Improved Vision-Language Alignment

DGX agent

arXiv:2606.07451v1 Announce Type: cross Abstract: Vision-language models such as CLIP are highly useful for diverse tasks due to their shared image-text embedding space. Despite this, the image and te

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

VeriDrive: Verifiable Counterfactual Supervision for Cost-Efficient Vision-Language Planning

DGX agent

arXiv:2606.07338v1 Announce Type: new Abstract: Vision-language driving models increasingly use reasoning supervision to bridge perception, prediction, and planning, but existing driving rationales ar

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

When Recovery Matters: The Blind Spot of Surrogate Privacy in MLLM Editing

DGX agent

arXiv:2606.07171v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) enable flexible instruction-driven image editing, but privacy risks arise when user images expose diverse and u

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

Beyond Output Matching: Preserving Internal Geometry in NVFP4 LLM Distillatio

DGX agent

arXiv:2606.05682v1 Announce Type: new Abstract: Demand for low-precision inference, including NVFP4-based approaches, has grown as large language models are increasingly deployed in latency and cost c

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Evaluating Agentic Configuration Repair for Computer Networks

DGX agent

arXiv:2606.06212v1 Announce Type: new Abstract: Misconfigurations in computer networks remain a major source of critical Internet outages. Research is turning to Large Language Models (LLMs) to automa

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Geographic Bias and Diversity in AI Evaluation

DGX agent

arXiv:2606.05187v1 Announce Type: cross Abstract: Among the many challenges hindering the responsible development and deployment of AI, arguably none has faced more intense scrutiny than bias in its v

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

GuardNet: Ensemble Strategies of Shallow Neural Networks for Robust Prompt Injection and Jailbreak Detection

DGX agent

arXiv:2606.05566v1 Announce Type: new Abstract: Large Language Models (LLMs) have transformed natural language processing, but they remain vulnerable to Prompt Injection (PI) and Jailbreak (JB) attack

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Multilingual Fine-Tuning via Localized Gradient Conflict Resolution

DGX agent

arXiv:2606.05613v1 Announce Type: new Abstract: The rapid evolution of Large Language Models (LLMs) has established cross-lingual versatility as a defining feature of modern systems. However, fine-tun

model-releasesarxiv-cs-ai
6 Jun 2026
Research

Representation Learning Enables Scalable Multitask Deep Reinforcement Learning

DGX agent

arXiv:2606.05555v1 Announce Type: cross Abstract: Scaling reinforcement learning (RL) to diverse multitask settings remains a central challenge. While recent advances in model-based RL achieve strong

researcharxiv-cs-ai
6 Jun 2026
Model Releases

SentinelBench: A Benchmark for Long-Running Monitoring Agents

DGX agent

arXiv:2606.05342v1 Announce Type: new Abstract: AI agents are increasingly asked to carry out work that spans minutes, hours, or longer. Yet the default model of agent behavior is continuous action: i

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Subspace-Aware Sparse Autoencoders for Effective Mechanistic Interpretability

DGX agent

arXiv:2606.06333v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) are widely used for mechanistic interpretability in large language models, yet their formulation assigns each latent featur

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

TokenMizer: Graph-Structured Session Memory for Long-Horizon LLM Context Management

DGX agent

arXiv:2606.06337v1 Announce Type: new Abstract: Large language model (LLM) deployments for long-horizon tasks face a fundamental constraint: context windows are finite while productive work sessions a

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

ToolChoiceConfusion: Causal Minimal Tool Filtering for Reliable LLM Agents

DGX agent

arXiv:2606.06284v1 Announce Type: new Abstract: Large language model agents increasingly rely on external tools, but larger tool menus can reduce reliability and efficiency by increasing wrong-tool ca

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Vortex: Efficient and Programmable Sparse Attention Serving for AI Agents

DGX agent

arXiv:2606.06453v1 Announce Type: new Abstract: Sparse attention is becoming increasingly important for serving large language models (LLMs) as generation lengths continue to grow. However, deploying

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

When Should Memory Stay Silent: Measuring Memory-Use Boundaries in Memory-Augmented Conversational Agents

DGX agent

arXiv:2606.06055v1 Announce Type: new Abstract: Long-term memory enables language model agents to support personalized interactions, but it remains unclear when available memories warrant integration

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

CLEAR: Cognition and Latent Evaluation for Adaptive Routing in End-to-End Autonomous Driving

DGX agent

arXiv:2606.06219v1 Announce Type: new Abstract: End-to-end autonomous driving models often struggle to balance multi-modal maneuver generation with real-time inference constraints. While diffusion mod

model-releasesarxiv-cs-ro
5 Jun 2026
Model Releases

CollabBench: Benchmarking and Unleashing Collaborative Ability of LLMs with Diverse Players via Proactive Engagement

DGX agent

arXiv:2606.05793v1 Announce Type: new Abstract: While LLM-based agents excel at individual tasks, effective collaboration with realistic human partners remains challenging. Most of the existing conver

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

DocHop-QA: Towards Multi-Hop Reasoning over Multimodal Document Collections

DGX agent

arXiv:2508.15851v2 Announce Type: replace Abstract: Despite rapid progress in large language models (LLMs), current QA benchmarks still overlook the core challenge of real-world scientific information

model-releasesarxiv-cs-cl
5 Jun 2026
Tutorials

Formal Concept Lattices are Good Semantic Scaffolds for Concept-Based Learning

DGX agent

arXiv:2606.05471v1 Announce Type: new Abstract: Learning semantics is essential for deep learning models to be interpretable and better aligned with human reasoning. Concept-based models approach this

tutorialsarxiv-cs-cv
5 Jun 2026
Model Releases

From Self to Other: Evaluating Demographic Perspective-Taking in LLM Hate Speech Annotation

DGX agent

arXiv:2606.06266v1 Announce Type: new Abstract: Hate speech detection is inherently subjective: people from different demographic groups perceive the same content very differently. Collecting enough a

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Generic Triple-Latent Compression with Gated Associative Retrieval

DGX agent

arXiv:2606.05175v1 Announce Type: new Abstract: We study generic triple-latent sequence models that maintain a running token state and compressed pair-memory pathway to capture higher-order token inte

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

IA-RAG: Interval-Algebra-Driven Temporal Reasoning for Dynamic Knowledge Retrieval

DGX agent

arXiv:2606.06044v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has shown strong effectiveness in grounding Large Language Models (LLMs) with external knowledge. However, existing

model-releasesarxiv-cs-cl
5 Jun 2026
Local Ai

Improving Heart-Focused Medical Question Answering in LLMs via Variance-Aware Rubric Rewards with GRPO

DGX agent

arXiv:2606.05174v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong promise in healthcare applications. Yet deploying general-purpose models in real-world settings remains d

local-aiarxiv-cs-cl
5 Jun 2026
Research

Learning to Route LLMs from Implicit Cost-Performance Preferences via Meta-Learning

DGX agent

arXiv:2606.06178v1 Announce Type: cross Abstract: Large language models (LLMs) present a trade-off between performance and cost, where more powerful models incur greater expense. LLM routing aims to m

researcharxiv-cs-cl
5 Jun 2026
Model Releases

LoRi: Low-Rank Distillation for Implicit Reasoning

DGX agent

arXiv:2606.05315v1 Announce Type: new Abstract: Implicit chain-of-thought (iCoT) methods aim to internalize reasoning in large language models, but often underperform explicit CoT prompting. We empiri

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Predict and Reconstruct: Joint Objectives for Self-Supervised Language Representation Learning

DGX agent

arXiv:2606.05173v1 Announce Type: new Abstract: Masked language modelling (MLM) has been the dominant pre-training objective for text encoders since BERT, yet it encourages representations that are st

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

ProSPy: A Profiling-Driven SQL-Python Agentic Framework for Enterprise Text-to-SQL

DGX agent

arXiv:2606.05836v1 Announce Type: new Abstract: Large language models have substantially advanced Text-to-SQL systems, yet applying them to enterprise-scale databases remains challenging. Real-world d

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Reducing Hallucinations in Complex Question Answering using Simple Graph-based Retrieval-Augmented Generation (long version)

DGX agent

arXiv:2606.05901v1 Announce Type: new Abstract: Large language models (LLMs) have fundamentally transformed the landscape of Natural Language Processing. Despite these advances, LLMs and LLM-based sys

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Self-supervised User Profile Generation for Personalization

DGX agent

arXiv:2606.05336v1 Announce Type: new Abstract: Personalizing large language models (LLMs) has become a central challenge as LLMs are deployed across recommendation, search, dialogue, and content gene

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

TextWand: A Unified Framework for Scene Text Editing

DGX agent

arXiv:2606.05730v1 Announce Type: new Abstract: We propose TextWand, a general-purpose framework that unifies scene text removal, generation, and replacement into a single model. By decomposing comple

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

The Mirage of Performance Gains: Why Contrastive Decoding Fails to Mitigate Object Hallucinations in MLLMs?

DGX agent

arXiv:2504.10020v4 Announce Type: replace Abstract: Contrastive decoding strategies are widely used to reduce object hallucinations in multimodal large language models (MLLMs). These methods work by c

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

When AI Says It Feels

DGX agent

arXiv:2606.05734v1 Announce Type: cross Abstract: Large language models (LLMs) are generally constrained from expressing feelings through human-preference alignment in post-training processes. This po

safetyarxiv-cs-cl
5 Jun 2026
Model Releases

YouZhi: Towards High-Concurrency Financial LLMs via Adaptive GQA-to-MLA Transition

DGX agent

arXiv:2606.05868v1 Announce Type: new Abstract: Large language models (LLMs) drive significant financial innovations, yet their high-concurrency deployment is severely bottlenecked by KV cache memory

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

100-LongBench: Are de facto Long-Context Benchmarks Literally Evaluating Long-Context Ability?

DGX agent

arXiv:2505.19293v2 Announce Type: replace-cross Abstract: Long-context capability is considered one of the most important abilities of LLMs, as a truly long context-capable LLM enables users to effort

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

A Systematic Evaluation of Positional Bias in Multi-Video Summarization with MLLMs

DGX agent

arXiv:2606.04596v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are increasingly used for video understanding, yet their reliability under multi-video inputs remains poorly un

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Activation-Based Active Learning for In-Context Learning: Challenges and Insights

DGX agent

arXiv:2606.05134v1 Announce Type: new Abstract: Deep active learning has previously been explored for LLM in-context sample selection, but not with methods that utilise recent advances in understandin

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Analysis-Driven Procedural Generation of an Engine Sound Dataset with Embedded Control Annotations

DGX agent

arXiv:2603.07584v2 Announce Type: replace-cross Abstract: Computational engine sound modeling is central to the automotive audio industry, particularly for active sound design applications and virtual

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Bayesian learning for the stochastic shortest path problem

DGX agent

arXiv:2606.04845v1 Announce Type: cross Abstract: Sequential decision-making problems are often modelled as a Markov decision process (MDP). We focus on the stochastic shortest path (SSP) problem, whi

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Building The Ph(ysical)AI Layer Of Machine Intelligence

DGX agent

arXiv:2606.04106v1 Announce Type: cross Abstract: Foundation models achieve generalization through massive-scale training on diverse data, but have limitations with transfer to truly unseen domains wi

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

DetectZoo: A Unified Toolkit for AI-Generated Content Detection Across Text, Audio, and Image Modalities

DGX agent

arXiv:2606.04205v1 Announce Type: cross Abstract: The growing popularity and capacity of generative models have eroded the distinction between human and machine-generated content, motivating a growing

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Dive into the Scene: Breaking the Perceptual Bottleneck in Vision-Language Decision Making via Focus Plan Generation

DGX agent

arXiv:2606.04046v1 Announce Type: cross Abstract: In embodied vision-language decision making tasks such as robotic manipulation and navigation, Vision-Language and Vision-Language-Action Models (VLMs

model-releasesarxiv-cs-ai
4 Jun 2026
Research

Evaluating Autoformalization Robustness via Semantically Similar Paraphrasing

DGX agent

arXiv:2511.12784v3 Announce Type: replace Abstract: Large Language Models (LLMs) have recently emerged as powerful tools for autoformalization. Despite their impressive performance, these models can s

researcharxiv-cs-cl
4 Jun 2026
Applications

Federated Learning for Multi-Center Sepsis Early Prediction with Privacy-Preserving

DGX agent

arXiv:2606.04338v1 Announce Type: new Abstract: Privacy-sensitive and distributed characteristics of multi-center medical data bring severe obstacles to centralized modeling for accurate early predict

applicationsarxiv-cs-lg
4 Jun 2026
Model Releases

FinTradeBench: A Financial Reasoning Benchmark for LLMs

DGX agent

arXiv:2603.19225v3 Announce Type: replace-cross Abstract: Real-world financial decision-making is a challenging problem that requires reasoning over heterogeneous signals, including company fundamenta

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Hyper-ICL: Attention Calibration with Hyperbolic Anchor Distillation for Multimodal In-Context Learning

DGX agent

arXiv:2606.04434v1 Announce Type: new Abstract: Multimodal In-Context Learning (ICL) has emerged as a practical inference paradigm for Multimodal Large Language Models, where a small set of interleave

model-releasesarxiv-cs-cv
4 Jun 2026
Tutorials

J-RAS: Mutual Adaptation for Medical Image Segmentation via Contrastive Retrieval-Augmented Joint Optimization

DGX agent

arXiv:2510.09953v3 Announce Type: replace Abstract: Manual medical image segmentation by clinicians, though accurate, is time-consuming and variable across experts, whereas AI-based models automate th

tutorialsarxiv-cs-cv
4 Jun 2026
← Previous
1…356357358359360…1067
Next →