AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Model Releases

AssumptionMiner: Extracting, Tracing, and Revising Implicit Assumptions in LLM Code Generation

DGX agent

arXiv:2607.22898v1 Announce Type: cross Abstract: Large language models (LLMs) generate code from natural-language prompts, yet real-world prompts rarely provide complete specifications. When prompts

model-releasesarxiv-cs-cl
28 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Between Suppression and Collapse: Evaluating Narrative Unlearning with LENS

DGX agent

arXiv:2607.22657v1 Announce Type: cross Abstract: Large language models (LLMs) can reproduce disinformation-aligned narrative frames as plausible explanations, raising the question of whether existing

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Compressing LLMs with MoP: Mixture of Pruners

DGX agent

arXiv:2602.06127v2 Announce Type: replace Abstract: The high computational demands of Large Language Models (LLMs) motivate methods that reduce parameter count and accelerate inference. In response, m

model-releasesarxiv-cs-lg
28 Jul 2026
Local Ai

Context-Aware Concept Distillation for Trustworthy Flood Prediction

DGX agent

arXiv:2607.23237v1 Announce Type: cross Abstract: Effective flood risk management relies on accurate forecasting, yet the 'black box' nature of stateof-the-art Deep Learning models creates a barrier t

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Cost-Aware Recovery-Pathway Identification and Bayesian Optimization for Autonomous Materials Discovery

DGX agent

arXiv:2607.23896v1 Announce Type: new Abstract: Autonomous laboratories automate experimental execution, but a campaign must also decide which recovery pathway merits optimization. We formulate this a

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Covariance-Boosted Gaussian Processes for Spatiotemporal Irregularities

DGX agent

arXiv:2607.23018v1 Announce Type: cross Abstract: Nonstationary Gaussian process (GP) models are powerful tools for capturing input-dependent variability by adapting to observed data. However, with li

safetyarxiv-cs-lg
28 Jul 2026
Safety

Data Pyramid for Embodied Manipulation

DGX agent

arXiv:2607.24744v1 Announce Type: cross Abstract: Multimodal foundation models learned to see and to speak by consuming the whole internet. Embodied agents admit no such shortcut, since they require d

safetyarxiv-cs-cv
28 Jul 2026
Model Releases

Do LLMs Know Their Vulnerable Scenarios?

DGX agent

arXiv:2607.23496v1 Announce Type: new Abstract: Safety-aligned large language models are trained to refuse harmful requests, yet embedding the same requests in particular scenarios can bypass their sa

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DocHRL: A Hierarchical Reinforcement Learning Framework for Cost-Optimised Document Classification

DGX agent

arXiv:2607.22644v1 Announce Type: new Abstract: Real-world document classification pipelines typically apply the same sequence of models to every incoming document, regardless of its complexity or typ

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DriveDNA: A Large-Scale Multimodal Naturalistic Driving Dataset and Benchmark for Driving Style Identification

DGX agent

arXiv:2607.23822v1 Announce Type: new Abstract: Driving style captures stable, driver-specific patterns in how a vehicle is driven. In naturalistic data, however, this signal is hard to isolate becaus

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

FilmBench: A Film-Grade Benchmark for Cinematic Video Generation

DGX agent

arXiv:2607.24241v1 Announce Type: cross Abstract: Progress in video generation keeps narrowing the visual gap between AI-generated and professionally produced footage, yet most benchmarks still draw p

model-releasesarxiv-cs-ai
28 Jul 2026
Research

LA-RL: Label-Aware Self-Reflection for Reinforcement Learning in Information Extraction

DGX agent

arXiv:2607.23420v1 Announce Type: new Abstract: Large language models show strong promise for information extraction (IE), but existing reflection-based correction methods are often misaligned with st

researcharxiv-cs-cl
28 Jul 2026
Model Releases

Language Shapes Instruction Hierarchy Compliance in Multilingual LLMs

DGX agent

arXiv:2607.23545v1 Announce Type: new Abstract: Instruction hierarchy (IH) requires models to prioritize instructions by source, ensuring that higher-priority instructions override lower-priority ones

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Mixture-of-Thought-Tokens: Unifying Perception and Reasoning for Free-form Multimodal Grounding

DGX agent

arXiv:2607.24407v1 Announce Type: new Abstract: Multimodal Large Language Models have made great progress in grounding tasks, yet existing methods still struggle to unify precise localization and comp

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Multi-Modal Scene Graph with Kolmogorov-Arnold Experts for Audio-Visual Question Answering

DGX agent

arXiv:2511.23304v2 Announce Type: replace Abstract: In this paper, we propose a novel Multi-Modal Scene Graph with Kolmogorov-Arnold Expert Network for Audio-Visual Question Answering (SHRIKE). The ta

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Neuromorphic Object Detection: An In-Depth Study and Future Directions

DGX agent

arXiv:2607.23576v1 Announce Type: new Abstract: Conventional frame-based cameras face significant challenges in detecting objects under high-speed motion blur or in low-light environments. Neuromorphi

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

No Optimal Language Set Exists for Multilingual Instruction Tuning: Insights from a Linguistically-Informed Study

DGX agent

arXiv:2410.07809v2 Announce Type: replace Abstract: Multilingual instruction tuning (MIT) is challenged by the curse of multilinguality, data scarcity, and high computational cost. A natural hypothesi

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

On the Impossibility of Unbiased and Length-Invariant Policy Optimization with Outcome Rewards

DGX agent

arXiv:2607.23364v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is the dominant reinforcement learning algorithm for training reasoning capabilities in large language models,

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

OpenAIs HealthBench in Action: Evaluating an LLM-Based Medical Assistant on Realistic Clinical Queries

DGX agent

arXiv:2509.02594v3 Announce Type: replace-cross Abstract: Evaluating large language models (LLMs) on their ability to generate high-quality, accurate, situationally aware answers to clinical questions

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Parameter-Efficient Adaptation of SAM3 for Prompt-Driven Surgical Concept Segmentation

DGX agent

arXiv:2607.23694v1 Announce Type: new Abstract: Efficient surgical segmentation empowers clinical diagnosis, intraoperative monitoring, and downstream robotic pipelines for reconstruction and simulati

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

ParBench: A Benchmark for Reliable Evaluation of LLM Parallel Code Translation

DGX agent

arXiv:2607.22588v1 Announce Type: new Abstract: Modern compute-intensive software must migrate across a changing ecosystem of accelerators, programming APIs, compiler stacks, and portability layers, i

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Phenology-based learning framework for yield estimation and harvest forecasting of raspberry fruits

DGX agent

arXiv:2411.00967v2 Announce Type: replace Abstract: The future of agriculture is intertwined with automation. Accurate fruit detection, yield estimation, and harvest time prediction are crucial for ef

model-releasesarxiv-cs-cv
28 Jul 2026
Local Ai

Poison to Detect: Detection of Targeted Overfitting in Federated Learning

DGX agent

arXiv:2509.11974v3 Announce Type: replace-cross Abstract: Federated Learning (FL) enables collaborative model training among clients without centralising data, making it a widely adopted privacy-enhan

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Poster: Rethinking Security in LLM Code Generation through Real-World Risk Scenarios

DGX agent

arXiv:2607.23088v1 Announce Type: cross Abstract: Large Language Models (LLMs) are widely used for code generation, yet their security behavior in realistic development workflows remains underexplored

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Random Forest-Based Prediction of Bone Volume Fraction and Fracture Position from S-Parameters

DGX agent

arXiv:2607.23563v1 Announce Type: new Abstract: In this paper, we propose a method for predicting bone volume fraction (BVF) and fracture position by constructing a random forest model based on multic

model-releasesarxiv-cs-lg
28 Jul 2026
Research

Same Question, Different Answers: Evaluating LLM Reliability Beyond Accuracy

DGX agent

arXiv:2607.22554v1 Announce Type: new Abstract: Large language models (LLMs) often achieve strong accuracy on benchmarks, yet it remains unclear how reliably they apply this knowledge when the same qu

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B

DGX agent

arXiv:2607.22545v1 Announce Type: cross Abstract: Deploying large language models in financial-services and agentic settings requires safety classifiers that simultaneously handle prompt injection, re

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Source-Free Controlled Adaptation of Teachers for Continual Test-Time Adaptation

DGX agent

arXiv:2607.23735v1 Announce Type: cross Abstract: In many real-world scenarios, encountering continual shifts in domain during inference is very common. Consequently, continual test-time adaptation (C

model-releasesarxiv-cs-cv
28 Jul 2026
Hardware

Spatial-IQ: Deconstructing Spatial Intelligence via Hierarchical Capability Tests

DGX agent

arXiv:2607.22864v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) excel at visual interpretation but fail on spatial reasoning tasks that humans solve reliably. Existing bench

hardwarearxiv-cs-ai
28 Jul 2026
Model Releases

TextRich: A Multi-Domain Benchmark for Detecting AI-Generated Text-Rich Images from GPT-Image-2

DGX agent

arXiv:2606.19259v2 Announce Type: replace-cross Abstract: Text-rich images often contain privacy-sensitive, transactional, or decision-relevant information. As recent multimodal image generation model

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

TokenMem: Faithful Knowledge Injection for Frozen LLMs

DGX agent

arXiv:2607.22625v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) enhances large language models (LLMs) with external knowledge, but suffers from knowledge conflicts: when retrieved

model-releasesarxiv-cs-ai
28 Jul 2026
Research

VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference

DGX agent

arXiv:2512.01031v2 Announce Type: replace-cross Abstract: Vision-Language-Action models (VLAs) are becoming increasingly capable across diverse robotic tasks. However, these models are typically deplo

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Weighted Low-Rank Matrix Approximation: Acceleration and Applications

DGX agent

arXiv:2109.11057v2 Announce Type: replace-cross Abstract: Weighted low-rank matrix approximation (WLRMA) generalizes classical low-rank approximation and matrix completion by allowing arbitrary elemen

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

XGRVFL-MV: Residual-Coupled Graph-Embedded Multi-View Random Vector Functional Link Network with FleXi Guardian Loss

DGX agent

arXiv:2607.23149v1 Announce Type: new Abstract: Random Vector Functional Link (RVFL) networks provide an efficient randomized learning framework for classification. Existing multi-view RVFL methods ut

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

Data Quality over Capacity: Internalizing Documents into LoRA Adapters for Closed-Book QA

DGX agent

arXiv:2607.21861v1 Announce Type: new Abstract: We study baking documents directly into the weights of a 4-bit Gemma-4-e4b model via LoRA, so a system can answer questions about a corpus closed-book:

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Do emulated quantum circuits change what CNNs look at? Performance and explainability comparison in medical image classification

DGX agent

arXiv:2607.21186v1 Announce Type: cross Abstract: Numerous studies have analyzed the use of hybrid quantum-classical convolutional neural networks as a promising alternative to classical deep learning

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Filling Before Advancing: Capability-Gap-Driven Post-Training for Scenario-Specialized Remote Sensing MLLMs

DGX agent

arXiv:2607.22205v1 Announce Type: new Abstract: Remote sensing multimodal large language models (RS-MLLMs) have improved general aerial-image understanding. However, Earth observation applications req

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Interpretable EEG biomarkers with bag-of-waves: Spatial and temporal waveform dictionaries for low-data regimes

DGX agent

arXiv:2607.22508v1 Announce Type: new Abstract: Electroencephalography (EEG) is widely used to diagnose neurological conditions, but its analysis usually relies on either predefined spectral features

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Learning What Matters: Supervising Sparse Attention Routing with Causal Evidence Sets

DGX agent

arXiv:2607.21692v1 Announce Type: cross Abstract: Sparse attention reduces the cost of long contexts by allowing each query to read only selected parts of the input. These selectors are often trained

model-releasesarxiv-cs-cl
27 Jul 2026
Research

Multi-Horizon Consistency as Geometry: When Latent Dynamics Contract, and When They Do Not

DGX agent

arXiv:2607.21645v1 Announce Type: new Abstract: Multi-horizon latent consistency is a common training knob in video predictors and world models, but practitioners rarely know what it does to transitio

researcharxiv-cs-lg
27 Jul 2026
Model Releases

Procedural Knowledge Is Not Low-Rank: Why LoRA Fails to Internalize Multi-Step Procedures

DGX agent

arXiv:2607.21612v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning methods like LoRA have become the default for adapting large language models, succeeding across instruction following,

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

SceneActBench: Can Agents Act on the 3D Scenes They See?

DGX agent

arXiv:2607.22393v1 Announce Type: cross Abstract: Vision-language model (VLM) agents increasingly use tools to act on 3D scenes rather than only describe them. Existing 3D benchmarks score textual res

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Spatially-Enhanced Temporal Fusion Transformer: Interpretable Multi-Output Prediction for Parametric Dynamical Systems with Time-Varying Inputs

DGX agent

arXiv:2505.00473v2 Announce Type: replace Abstract: We explore the promising performance of a transformer model in predicting outputs of parametric dynamical systems with external time-varying input s

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Variational Low-rank Tensor Decomposition for Multisubject Spatiotemporal Data Analysis

DGX agent

arXiv:2607.22262v1 Announce Type: cross Abstract: Modeling shared and subject-specific structure in multisubject spatiotemporal data remains challenging, particularly in neuroimaging, where both spati

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Faster IndexTTS-2: Accelerating and Streaming Autoregressive Zero-Shot Text-to-Speech Synthesis on GPUs

DGX agent

arXiv:2607.21042v1 Announce Type: new Abstract: Autoregressive text-to-speech models achieve strong naturalness but suffer from slow inference due to sequential token generation, limiting their deploy

model-releasesarxiv-cs-ai
24 Jul 2026
Tutorials

Gumbel Distillation for Parallel Text Generation

DGX agent

arXiv:2603.22216v2 Announce Type: replace Abstract: The slow, sequential nature of autoregressive (AR) language models has driven the adoption of parallel decoding methods. However, these non-AR model

tutorialsarxiv-cs-cl
24 Jul 2026
Research

Interpretable Embeddings with Sparse Autoencoders: A Data Analysis Toolkit

DGX agent

arXiv:2512.10092v2 Announce Type: replace Abstract: Analyzing large-scale text corpora is a core challenge in machine learning, crucial for tasks like identifying undesirable model behaviors or biases

researcharxiv-cs-ai
24 Jul 2026
Research

Knowledge Injection Exists in MoE? Exploring Expert-Aware Contrast Decoding in MoE for Mitigating LLMs'Hallucinations

DGX agent

arXiv:2607.20426v1 Announce Type: cross Abstract: Existing LLM hallucination mitigation methods, including prompt engineering and model optimization, either hardly alter models'internal knowledge or h

researcharxiv-cs-ai
24 Jul 2026
← Previous
1…308309310311312…1065
Next →