AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,191 results
Model Releases

Integrating Deep Learning Demand Forecasting with Multi-Objective Optimization for Circular Coffee Supply Chains: A Data-Driven Framework for Cost, Emissions, and Freshness Management

DGX agent

arXiv:2606.08314v1 Announce Type: new Abstract: The coffee supply chain is one of the most complex agri-food networks, marked by geographically dispersed production, multi-tier coordination, and high

model-releasesarxiv-cs-ai
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Integrating gene regulatory priors into Transformer attention with scTransformer for interpretable scRNA-seq analysis

DGX agent

arXiv:2606.09558v1 Announce Type: cross Abstract: Motivation: Transformer-based models are increasingly applied to large-scale single-cell transcriptomics, showing strong performance through self-supe

researcharxiv-cs-lg
9 Jun 2026
Model Releases

Internalizing Geometric Law: Learning from Solver Residuals for Precision-Critical Generation

DGX agent

arXiv:2606.09278v1 Announce Type: cross Abstract: Large Language Models frequently hallucinate in precision-critical domains such as technical diagramming and mechanical design, where outputs must sat

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

KPGrasp: Scalable Keypoint Flow Matching for Dexterous Grasp Generation

DGX agent

arXiv:2606.09314v1 Announce Type: new Abstract: Generating high-quality dexterous grasps remains challenging for learning-based methods, which often depend on carefully tuned contact losses or costly

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

Learning Predictive Control with Deep Koopman Operators for Autonomous Vehicle Motion Planning

DGX agent

arXiv:2606.08136v1 Announce Type: new Abstract: Model Predictive Control (MPC) is widely used for autonomous-vehicle (AV) motion planning, but its real-time applicability is often limited by the need

model-releasesarxiv-cs-ro
9 Jun 2026
Research

Learning to Solve Generative ODEs Beyond the Linear Span

DGX agent

arXiv:2606.08672v1 Announce Type: new Abstract: Diffusion and flow generative models sample by integrating a learned ODE, but high quality still requires many sequential model evaluations. Solver lear

researcharxiv-cs-cv
9 Jun 2026
Model Releases

LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load

DGX agent

arXiv:2603.23640v2 Announce Type: replace-cross Abstract: Deploying large language models on-device for always-on personal agents demands sustained inference from hardware tightly constrained in power

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

MEnvAgent: Scalable Polyglot Environment Construction for Verifiable Software Engineering

DGX agent

arXiv:2601.22859v3 Announce Type: replace-cross Abstract: The evolution of Large Language Model (LLM) agents for software engineering (SWE) is constrained by the scarcity of verifiable datasets, a bot

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Minibatch Selection via Partition Matroid Constrained Gradient Matching

DGX agent

arXiv:2606.07954v1 Announce Type: cross Abstract: Training large language models (LLMs) on heterogeneous data requires selecting minibatches that balance convergence speed with coverage across domains

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

OmniGen-AR: AutoRegressive Any-to-Image Generation

DGX agent

arXiv:2606.09156v1 Announce Type: new Abstract: Autoregressive (AR) models have demonstrated strong potential in visual generation, offering superior performance with simple architectures and optimiza

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

OmniMem: Perturbation-aware Memory Compression for Streaming Audio-Visual LLMs

DGX agent

arXiv:2606.07577v1 Announce Type: new Abstract: Audio-visual large language models (LLMs) hold strong promise for long-form video understanding, yet their long-video inference is fundamentally limited

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Partially Performative Prediction

DGX agent

arXiv:2606.07890v1 Announce Type: new Abstract: Performative prediction studies feedback loops that arise when predictive models are deployed in consequential domains. In these settings, deploying a m

researcharxiv-cs-lg
9 Jun 2026
Model Releases

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs

DGX agent

arXiv:2606.09038v1 Announce Type: new Abstract: Large Language Models (LLMs) have enabled increasingly personalized interactions by adapting to users' preferences, contexts, and long-term histories. H

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

phepy: Visual benchmarks and improvements for out-of-distribution detectors

DGX agent

arXiv:2503.05169v2 Announce Type: replace Abstract: Applying machine learning to increasingly high-dimensional problems with sparse or biased training data increases the risk that a model is used on i

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Pre-Intervention Prediction of Sparse Autoencoder Steering Side Effects

DGX agent

arXiv:2606.08365v1 Announce Type: cross Abstract: Sparse autoencoder (SAE) features are increasingly used to steer language models, but feature steering is rarely clean: the same intervention can beha

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Revisiting Training Scale: An Empirical Study of Token Count, Power Consumption, and Parameter Efficiency

DGX agent

arXiv:2601.06649v2 Announce Type: replace-cross Abstract: Research in machine learning has questioned whether increases in training token counts reliably produce proportional performance gains in larg

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SC3: The Multi-Solvent Solubility Challenge and Benchmark

DGX agent

arXiv:2606.07656v1 Announce Type: cross Abstract: Solubility prediction is a standard benchmark in computational chemistry, yet multi-solvent models which reportedly approach the experimental-noise ce

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

SciFlow-Bench: Evaluating Structure-Aware Scientific Diagram Generation via Inverse Parsing

DGX agent

arXiv:2602.09809v2 Announce Type: replace Abstract: Scientific diagrams convey explicit structural information, yet modern text-to-image models often produce visually plausible but structurally incorr

model-releasesarxiv-cs-cv
9 Jun 2026
Local Ai

Self-Consistent Generative Paths via Admissible Random Variational Transport

DGX agent

arXiv:2606.08953v1 Announce Type: new Abstract: Modern generative models often define an entire probability path from a simple prior to the data law, rather than only an endpoint map. Diffusion models

local-aiarxiv-cs-lg
9 Jun 2026
Model Releases

Shared Latent Structures Enable Unified Backdoor Detection and Mitigation in LLMs

DGX agent

arXiv:2606.07963v1 Announce Type: new Abstract: Backdoor attacks in large language models (LLMs) are often treated as isolated trigger-response failures, motivating defenses tailored to specific trigg

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Shared Semantics, Divergent Mechanisms: Unsupervised Feature Discovery by Aligning Semantics and Mechanisms

DGX agent

arXiv:2606.08236v1 Announce Type: cross Abstract: As large language models are increasingly deployed in high-stakes settings, there is a growing need for tools that audit not only model outputs but al

researcharxiv-cs-lg
9 Jun 2026
Model Releases

SNN-MLIR: An MLIR Dialect for Compiling Neuromorphic SNNs from NIR to Bare-Metal C

DGX agent

arXiv:2606.09213v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) are increasingly trained in a wide range of frameworks (SnnTorch, Lava, Norse, and others) each with its own model form

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving

DGX agent

arXiv:2606.08635v1 Announce Type: new Abstract: Prefill-decode (PD) disaggregation decouples prompt processing from token generation, but it also turns the key-value (KV) cache into a network payload.

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

Steganography Without Modification: Hidden Communication via LLM Seeds

DGX agent

arXiv:2606.09135v1 Announce Type: cross Abstract: We demonstrate that widely deployed Large Language Model (LLM) inference stacks harbor a steganographic channel that requires no modification to model

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Struct-Searcher: Agentic Structural Thinking Advances Multimodal Deep Information Seeking

DGX agent

arXiv:2606.07689v1 Announce Type: new Abstract: Deep research agents have attracted increasing attention for their ability to collect large-scale online information to acquire target knowledge, with r

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

The Injection Paradox: Brand-Level Suppression in Safety-Trained LLM Recommendations via RAG Context Injection

DGX agent

arXiv:2606.09204v1 Announce Type: new Abstract: We present a reproducible failure mode of safety training in RAG-based LLM recommendation -- the Injection Paradox -- in which prompt injections embedde

model-releasesarxiv-cs-lg
9 Jun 2026
Research

The Mirrored Influence Hypothesis: Efficient Data Influence Estimation by Harnessing Forward Passes

DGX agent

arXiv:2402.08922v3 Announce Type: replace Abstract: Large-scale black-box models have become ubiquitous across numerous applications. Understanding the influence of individual training data sources on

researcharxiv-cs-lg
9 Jun 2026
Model Releases

TQA-Bench: Evaluating LLMs for Multi-Table Question Answering

DGX agent

arXiv:2411.19504v2 Announce Type: replace Abstract: The advance of large language models (LLMs) has unlocked great opportunities in complex multi-modal data management tasks, particularly in question

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

TRL-Bench: Standardizing Cross-Paradigm Representation-Level Evaluation of Tabular Encoders

DGX agent

arXiv:2606.09323v1 Announce Type: new Abstract: Tabular encoders are usually evaluated inside task-specific end-to-end pipelines, so models from different training paradigms are difficult to compare d

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Unification of Closed-Open Industrial Detection Scenarios: New Large-Scale Benchmarks,Challenges and Baselines

DGX agent

arXiv:2606.07953v1 Announce Type: new Abstract: Large-scale Visual-Language Models (LVLMs) have achieved remarkable success in natural visual tasks, yet their application to industrial defect detectio

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

UniQL: Towards Dialect-Universal Benchmarking for Text-to-SQL

DGX agent

arXiv:2606.08018v1 Announce Type: new Abstract: Existing text-to-SQL benchmarks are largely centered on SQLite, making it difficult to evaluate whether models can generalize across heterogeneous SQL d

model-releasesarxiv-cs-ai
9 Jun 2026
Research

VFEM: Visual Feature Empowered Multivariate Time Series Forecasting with Cross-Modal Fusion

DGX agent

arXiv:2510.03244v2 Announce Type: replace-cross Abstract: Large time series foundation models often adopt channel-independent architectures to handle varying data dimensions, but this design ignores c

researcharxiv-cs-ai
9 Jun 2026
Model Releases

A Geometric Gaussian Mixture Representation of Plane Curves

DGX agent

arXiv:2606.06505v1 Announce Type: cross Abstract: We introduce a user defined probabilistic polygonal representation for plane curves. Given a curve, we select vertices on the curve and connect consec

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Act As a Real Researcher: A Suite of Benchmarks Evaluating Frontier LLMs and Agentic Harnesses in Research Lifecycle

DGX agent

arXiv:2606.07462v1 Announce Type: new Abstract: As foundation models advance and agent scaffolding becomes increasingly sophisticated, agents have demonstrated remarkable proficiency in complex, long-

model-releasesarxiv-cs-ai
8 Jun 2026
Research

From Sampled Outcomes to Capability Distributions: Rethinking Supervision for LLM Routing

DGX agent

arXiv:2606.06924v1 Announce Type: new Abstract: Existing LLM routing methods typically treat a model's single response to a query as its capability label for training routers. However, because LLM gen

researcharxiv-cs-lg
8 Jun 2026
Research

Generative Molecular Morphing for Flexible-Size Design via Unbalanced Optimal Transport

DGX agent

arXiv:2606.07239v1 Announce Type: new Abstract: The success of generative molecular design hinges on a model's steerability toward high-reward samples. Because many molecular properties are intrinsica

researcharxiv-cs-lg
8 Jun 2026
Model Releases

Hierarchical Certified Semantic Commitment for Byzantine-Resilient LLM-Agent Collaboration

DGX agent

arXiv:2606.07316v1 Announce Type: cross Abstract: Byzantine collaboration among large-language-model agents requires a finality-control primitive: given delivered stochastic, structured natural-langua

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Hierarchical Semantic-Constrained Heterogeneous Graph for Audio-Visual Event Localization

DGX agent

arXiv:2606.07033v1 Announce Type: new Abstract: Open-vocabulary audio-visual event localization (OV-AVEL) jointly models audio-visual cues to recognize and temporally localize events, including catego

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Improving Cross-Lingual Factual Recall via Consistency-Driven Reinforcement Learning

DGX agent

arXiv:2606.06586v1 Announce Type: new Abstract: Large language models (LLMs) trained predominantly on English data encode substantial world knowledge, yet often fail to express it reliably in other la

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

It's a TRAP! Task-Redirecting Agent Persuasion Benchmark for Web Agents

DGX agent

arXiv:2512.23128v2 Announce Type: replace-cross Abstract: Web-based agents powered by large language models are increasingly used for tasks such as email management or professional networking. Their r

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

MemDreamer: Decoupling Perception and Reasoning for Long Video Understanding via Hierarchical Graph Memory and Agentic Retrieval Mechanism

DGX agent

arXiv:2606.07512v1 Announce Type: cross Abstract: Current Vision-Language Models struggle with hours-long videos because processing full-length visual sequences induces prohibitive token explosion and

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

Online Pandora's Box for Contextual LLM Cascading

DGX agent

arXiv:2606.07392v1 Announce Type: new Abstract: Motivated by Large Language Model (LLM) cascading, we propose an online contextual Pandora's Box model for adaptively querying and selecting LLM APIs. I

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

PromptPrint: Behavioral Biometrics Through Natural Language Prompting in LLMs

DGX agent

arXiv:2606.06755v1 Announce Type: new Abstract: Authorship attribution research has traditionally focused on long-form, expressive texts; however, interactions with large language models (LLMs) are ty

model-releasesarxiv-cs-cl
8 Jun 2026
Research

Real-Time AttentionBender: Granular Interactive Network Bending of Video Diffusion Transformers

DGX agent

arXiv:2606.06497v1 Announce Type: cross Abstract: Generative video models have achieved remarkable visual fidelity, yet their prompt-only interface offers thin creative agency and obscures the model's

researcharxiv-cs-cv
8 Jun 2026
Model Releases

REMEDI: A Benchmark for Retention and Unlearning Evaluation in Multi-label Clinical Disease Inference

DGX agent

arXiv:2606.07141v1 Announce Type: cross Abstract: Language models trained for clinical disease inference are trained on patient data, which may include sensitive and private information, and data owne

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

RETROSPECT: RETROsynthesis via Sequential Prediction, and Chemically Transformed-ranking

DGX agent

arXiv:2606.07181v1 Announce Type: cross Abstract: Single-step retrosynthesis needs both accurate first-ranked suggestions and candidate lists that are rich enough for downstream selection. We study th

model-releasesarxiv-cs-ai
8 Jun 2026
Hardware

Scalable Joint Resource Allocation for SLO-Constrained LLM Inference in Heterogeneous GPU Clouds

DGX agent

arXiv:2604.07472v2 Announce Type: replace Abstract: Serving large language model (LLM) inference in cloud environments requires jointly optimizing model selection, GPU provisioning, parallelism config

hardwarearxiv-cs-lg
8 Jun 2026
Model Releases

Seeing Without Exposing: Adaptive Privacy Control for Open-World, Context-Hungry MLLMs

DGX agent

arXiv:2606.07175v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have raised new privacy challenges. On the data side, user-provided inputs often include unpredictable sensitiv

model-releasesarxiv-cs-cv
8 Jun 2026
← Previous
1…355356357358359…1067
Next →