AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,935 results
Model Releases

Enabling Progressive Whole-slide Image Analysis with Multi-scale Pyramidal Network

DGX agent

arXiv:2602.01951v2 Announce Type: replace Abstract: Multiple-instance Learning (MIL) is commonly used for computational pathology (CPath), where multi-scale features are essential for capturing both f

model-releasesarxiv-cs-cv
10 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Fisher-Guided Progressive Parameter Selection for Adaptive Fine-Tuning

DGX agent

arXiv:2606.10196v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) aims to adapt pretrained models with a small trainable parameter subset, however, most existing methods choose

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Flaws in the LLM Automation Narrative

DGX agent

arXiv:2606.11166v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly described as performing at the level of human experts on knowledge economy tasks. These claims are prima

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

GhazalBench: Evaluating LLM Understanding and Canonical Surface-Form Access in Persian Ghazals

DGX agent

arXiv:2603.09979v2 Announce Type: replace Abstract: Persian poetry plays an active role in Iranian cultural practice, where verses by canonical poets such as Hafez are frequently quoted, paraphrased,

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

GRID: Scaling Task-Agnostic Inference in Continual Prompt Tuning

DGX agent

arXiv:2507.14725v4 Announce Type: replace-cross Abstract: Prompt-based continual learning (CL) offers a parameter-efficient way to adapt large language models (LLMs) across task sequences. However, ex

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

LakeQA: An Exploratory QA Benchmark over a Million-Scale Data Lake

DGX agent

arXiv:2606.10460v1 Announce Type: cross Abstract: Recent large language models (LLMs) have shown rapid progress in reading-based question answering (QA), where evidence is explicitly provided or can b

model-releasesarxiv-cs-ai
10 Jun 2026
Research

Machine Learning Methods for Studying Latent Neural Activity Dynamics

DGX agent

arXiv:2606.10530v1 Announce Type: cross Abstract: Recent developments in brain recording are driving a demand for machine learning tools capable of decoding the latent structure of large populations o

researcharxiv-cs-ai
10 Jun 2026
Model Releases

MemVenom: Triggered Poisoning of Multimodal Memories in Web Agents

DGX agent

arXiv:2606.10742v1 Announce Type: cross Abstract: External memory has become a core component of modern web agents, enabling long-horizon reasoning through the retrieval of past experiences. However,

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Mixtures of Neural Operators Reduce Active Complexity in Operator Learning

DGX agent

arXiv:2404.09101v3 Announce Type: replace-cross Abstract: Operator-learning systems are not governed solely by total parameter count; for one query, the relevant bottleneck can be the model that must

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Nonlinear Estimator: Dual Bayesian Affine Estimators for Parameter Learning

DGX agent

arXiv:2606.10111v1 Announce Type: new Abstract: This paper presents a nonlinear parameter estimator for Wiener-type state-space models obtained as a fixed-point architecture that couples two affine mi

model-releasesarxiv-cs-lg
10 Jun 2026
Hardware

One-Step Residual Shifting Diffusion for Image Super-Resolution via Distillation

DGX agent

arXiv:2503.13358v5 Announce Type: replace Abstract: Diffusion models for super-resolution (SR) produce high-quality visual results but require expensive computational costs. Despite the development of

hardwarearxiv-cs-cv
10 Jun 2026
Model Releases

Online Self-Training for Co-Adaptation in Hierarchical Diffusion Policies

DGX agent

arXiv:2603.05291v2 Announce Type: replace Abstract: Hierarchical policies decompose language-conditioned long-horizon robotic manipulation into a high-level planner and a low-level controller. However

model-releasesarxiv-cs-ro
10 Jun 2026
Applications

Privacy-Preserving Credit Risk Prediction with Alternative Data

DGX agent

arXiv:2606.10333v1 Announce Type: new Abstract: Credit risk prediction is a critical problem in the consumer credit industry. Traditionally, financial institutions construct credit risk prediction mod

applicationsarxiv-cs-lg
10 Jun 2026
Model Releases

Quality Is Not a Safety Proxy Under Quantization

DGX agent

arXiv:2606.10154v1 Announce Type: new Abstract: Quantized checkpoints are often screened first with quality metrics and only later, if at all, with direct safety tests. This paper audits that shortcut

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

REAL: A Reasoning-Enhanced Graph Framework for Long-Term Memory Management of LLMs

DGX agent

arXiv:2606.10694v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly expected to interact with users over long time horizons. However, due to their finite context window, LLMs

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI

DGX agent

arXiv:2605.06234v2 Announce Type: replace Abstract: Embodied AI is a prominent research topic in both academia and industry. Current research centers on completing tasks based on explicit user instruc

model-releasesarxiv-cs-ro
10 Jun 2026
Research

SPACR: Single-Pass Adaptive Training of Uncertainty-Aware Conformal Regressors

DGX agent

arXiv:2606.10734v1 Announce Type: new Abstract: Conformal Prediction (CP) provides robust uncertainty guarantees for predictive models, but is typically applied post hoc, which misaligns model trainin

researcharxiv-cs-lg
10 Jun 2026
Safety

Stop Early, Spend Less: Hidden-State Probes as a Practical Recipe for Streaming Moderation of LLM Outputs

DGX agent

arXiv:2606.10487v1 Announce Type: cross Abstract: Deploying large language models in user-facing systems requires efficient output safety filtering. Existing approaches typically rely on a separate mo

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

Structure from Reasoning, Numbers from Search: On-Premise Open LLMs as Structural Priors for Coupled MIMO Controller Tuning

DGX agent

arXiv:2606.11015v1 Announce Type: new Abstract: Tuning controllers for strongly coupled multi-input multi-output (MIMO) industrial processes is hard: decentralized classical auto-tuning ignores loop i

model-releasesarxiv-cs-ai
10 Jun 2026
Applications

Supervised Fine-tuning with Synthetic Rationale Data Hurts Real-World Disease Prediction

DGX agent

arXiv:2606.10279v1 Announce Type: new Abstract: Supervised fine-tuning with synthetic rationale data is widely assumed to improve language model performance on clinical prediction tasks by teaching mo

applicationsarxiv-cs-ai
10 Jun 2026
Model Releases

Temporal Sheaf Neural Networks with Dynamic Orthogonal Transport

DGX agent

arXiv:2606.10071v1 Announce Type: cross Abstract: We introduce Temporal Sheaf Neural Networks (TSNN), a temporal link prediction framework that equips each node with a time-varying orthogonal frame an

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

The Role of Feedback Alignment in Self-Distillation

DGX agent

arXiv:2606.11173v1 Announce Type: new Abstract: Conditioning a language model on additional context, such as feedback on a previous attempt, typically improves its response. Self-distillation trains t

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

Trainable Smooth-Rotation Transforms with Learned Channel Scales for LLM Quantization

DGX agent

arXiv:2606.09927v1 Announce Type: cross Abstract: Post-training quantization (PTQ) is one of the most practical ways to reduce the serving cost of Large Language Models (LLMs), but activation quantiza

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Training LLMs to Enforce Multi-Level Instruction Hierarchies via Gravity-Weighted Direct Preference Optimization

DGX agent

arXiv:2606.10860v1 Announce Type: cross Abstract: Production LLMs receive instructions from sources with very different levels of trust, yet attend to every token with uniform architectural privilege.

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

V-REX: Benchmarking Exploratory Visual Reasoning via Chain-of-Questions

DGX agent

arXiv:2512.11995v2 Announce Type: replace-cross Abstract: While many vision-language models (VLMs) are developed to answer well-defined, straightforward questions with highly specified targets, as in

model-releasesarxiv-cs-ai
10 Jun 2026
Research

Whisfusion: Parallel ASR Decoding with Masked Diffusion

DGX agent

arXiv:2508.07048v2 Announce Type: replace-cross Abstract: Autoregressive (AR) encoder-decoder models dominate high-quality multilingual ASR, but their left-to-right decoders make inference latency sca

researcharxiv-cs-ai
10 Jun 2026
Model Releases

XtrAIn: Training-Guided Occlusion for Feature Attribution

DGX agent

arXiv:2606.10877v1 Announce Type: cross Abstract: Occlusion-based attribution methods provide an intuitive way to estimate feature importance by perturbing input features and measuring the resulting c

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

A Framework for Evaluating and Benchmarking Concept Drift Detection Methods

DGX agent

arXiv:2606.07789v1 Announce Type: new Abstract: Data stream mining is fundamentally challenged by concept drift, where distributional changes can degrade model performance. Despite the proliferation o

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

A Unifying Framework for Concept-Based Representational Similarity

DGX agent

arXiv:2606.09653v1 Announce Type: new Abstract: Learned representations across models and modalities often exhibit striking structural similarities, suggesting shared underlying concept decompositions

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

AgentCompile: An LLM-Guided Compiler for Direct CUDA Inference

DGX agent

arXiv:2606.07665v1 Announce Type: cross Abstract: Transformer inference increasingly depends on specialized compiler and runtime support, but real model graphs still require semantic decisions about w

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

AI Scientists Are Only as Good as Their Evidence: A Stratified Ablation of Proprietary Data and Reasoning Skills in Drug-Asset Valuation

DGX agent

arXiv:2606.09556v1 Announce Type: new Abstract: AI Scientist agents are often evaluated as if capability were mainly a function of model quality, prompting, or reasoning scaffolds. We test a different

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Benchmarking Open-Ended Multi-Agent Coordination in Language Agents

DGX agent

arXiv:2606.08340v1 Announce Type: new Abstract: As language models are increasingly deployed as autonomous agents, they must coordinate with others over long horizons in open-ended interactive tasks.

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Beyond Consistency: Preserving Temporal Structure in Zero-Shot Video Editing

DGX agent

arXiv:2606.08780v1 Announce Type: new Abstract: Existing zero-shot video editing methods rely on pre-trained diffusion models, successfully achieving spatial control and basic temporal consistency but

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Beyond Pass/Fail: Using Process Mining to Understand How LLMs Resist (and Fail) Red Team Attacks

DGX agent

arXiv:2606.07833v1 Announce Type: cross Abstract: Standard AI red teaming evaluations reduce adversarial campaigns to a single binary outcome, attack success rate (ASR), not taking into account the se

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Bidirectional Semantic Complementary Tool Retrieval for Remote Sensing Agents

DGX agent

arXiv:2606.07538v1 Announce Type: cross Abstract: Large language model (LLM)-based agents provide a novel paradigm for the automated processing of remote sensing(RS) data. Their success in complex RS

model-releasesarxiv-cs-ai
9 Jun 2026
Research

CapRL++: Unified Reinforcement Learning with Verifiable Rewards for Dense Image and Video Captioning

DGX agent

arXiv:2606.09393v1 Announce Type: new Abstract: Image and video captioning are fundamental tasks that bridge the visual and linguistic domains, playing a critical role in pre-training Large Vision-Lan

researcharxiv-cs-cv
9 Jun 2026
Model Releases

CHROMA: Detecting AI-Generated Images through Inter-Channel Color-Space Correlations

DGX agent

arXiv:2606.08864v1 Announce Type: new Abstract: The rapid adoption of diffusion and large-scale generative models has made it increasingly challenging to distinguish synthetic imagery from real photog

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

ComplexConstraints and Beyond: Expert Rubrics for RLVR

DGX agent

arXiv:2606.09118v1 Announce Type: new Abstract: As LLM capabilities advance rapidly, the evaluation methods used to assess them increasingly lag behind. Traditional benchmarks relied on programmatic v

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

CURE: Curriculum-guided Multi-task Training for Reliable Anatomy Grounded Report Generation

DGX agent

arXiv:2601.15408v2 Announce Type: replace-cross Abstract: Medical vision-language models can automate the generation of radiology reports but struggle with accurate visual grounding and factual consis

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Deep Tree Tensor Networks

DGX agent

arXiv:2502.09928v2 Announce Type: replace-cross Abstract: Originating in quantum physics, tensor networks (TNs) have been widely adopted as exponential machines and parametric decomposers for recognit

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Disjoint Generation of Synthetic Data

DGX agent

arXiv:2507.19700v2 Announce Type: replace Abstract: We propose a new framework for generating tabular synthetic datasets via disjoint generative models. In this paradigm, a dataset is partitioned into

researcharxiv-cs-lg
9 Jun 2026
Model Releases

Frequency-Domain Latent Attention Gating for Cross-Domain Token Aggregation

DGX agent

arXiv:2606.08191v1 Announce Type: cross Abstract: Token aggregation is a common bottleneck in models that map token representations to sample-level predictions, yet most pooling methods operate only i

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Graph2Idea:Retrieval-Augmented Scientific Idea Generation with Graph-Structured Contexts

DGX agent

arXiv:2606.09105v1 Announce Type: new Abstract: Generating novel, feasible, and high-quality research ideas is an important yet challenging task in scientific discovery.Recent Large Language Model (LL

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

HDSL: A Hierarchical Domain-Specific Language for Structured 3D Indoor Scene Generation and Localized Editing with LLM Agents

DGX agent

arXiv:2606.09738v1 Announce Type: new Abstract: Text-driven indoor scene generation and editing require an intermediate representation that language models can both produce and revise. Existing LLM-ba

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

iOSWorld: A Benchmark for Personally Intelligent Phone Agents

DGX agent

arXiv:2606.09764v1 Announce Type: new Abstract: A useful phone agent needs to be personally intelligent. It should reason over a user's identity, history, and preferences as they exist on the device,

model-releasesarxiv-cs-lg
9 Jun 2026
Hardware

Kunlun: Establishing Scaling Laws for Massive-Scale Recommendation Systems through Unified Architecture Design

DGX agent

arXiv:2602.10016v3 Announce Type: replace-cross Abstract: Deriving predictable scaling laws that govern the relationship between model performance and computational investment is crucial for designing

hardwarearxiv-cs-ai
9 Jun 2026
Model Releases

Language as a Sensor: Calibrated Spatial Belief Estimation in 3D Scenes from Natural Language

DGX agent

arXiv:2606.08666v1 Announce Type: new Abstract: Robots deployed in human-centric environments routinely receive natural-language descriptions of spatial information ('I left my backpack on the table')

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

MedVision: Benchmarking Quantitative Medical Image Analysis

DGX agent

arXiv:2511.18676v2 Announce Type: replace-cross Abstract: Current vision-language models (VLMs) in medicine are primarily designed for categorical question answering (e.g., 'Is this normal or abnormal

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…414415416417418…1082
Next →