AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Model Releases

PennySynth: RAG-Driven Data Synthesis for Automated Quantum Code Generation

DGX agent

arXiv:2605.25572v1 Announce Type: cross Abstract: The growing complexity of quantum programming frameworks has exposed a critical limitation in existing large language model (LLM)-based code assistant

model-releasesarxiv-cs-ai
26 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Reward-free Alignment for Conflicting Objectives

DGX agent

arXiv:2602.02495v3 Announce Type: replace-cross Abstract: Direct alignment methods are increasingly used to align large language models (LLMs) with human preferences. However, many real-world alignmen

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SafeCtrl-RL: Inference-Time Adaptive Behaviour Control for LLM Dialogue via RL-Driven Prompt Optimisation

DGX agent

arXiv:2605.25984v1 Announce Type: cross Abstract: Ensuring safe and contextually appropriate behaviour in Large Language Models (LLMs) remains a critical challenge for real-world deployment. We presen

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SemanticZip: A Pilot Framework for Lossy Text Compression with LLMs as Semantic Decompressors

DGX agent

arXiv:2605.24541v1 Announce Type: cross Abstract: Text compression for large language model (LLM) systems is usually framed as token deletion, retrieval, summarization, or exact reconstruction. We stu

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Stochastic Linear Bandits with Parameter Noise

DGX agent

arXiv:2601.23164v2 Announce Type: replace Abstract: We study the stochastic linear bandits with parameter noise model, in which the reward of action a is a^op heta where heta is sampled i.i.d. We show

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

STREAM: A Data-Centric Framework for Mining High-Value Task-Oriented Dialogues from Streaming Media

DGX agent

arXiv:2605.25162v1 Announce Type: cross Abstract: Large language models for vertical domains are bottlenecked by the scarcity of complex, domain-specific task-oriented dialogues. Existing data acquisi

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

StreamProfileBench: A Benchmark for Fine-Grained User Profile Inference in Real-World Streaming Scenarios

DGX agent

arXiv:2605.25758v1 Announce Type: new Abstract: Large Language Models (LLMs) have reshaped user profiling, yet current evaluations mainly focus on static data snapshots. This paradigm overlooks the re

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

The Perception-Physics Paradox: Probing Scientific Alignment with TC-Bench

DGX agent

arXiv:2605.24782v1 Announce Type: new Abstract: While Vision Foundation Models (VFMs) excel at predictive tasks on satellite imagery, their performance can arise from visual correlations rather than u

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning

DGX agent

arXiv:2605.25850v1 Announce Type: cross Abstract: This paper investigates large language model (LLM) abstention learning, specifically using ternary reward, which incentivize truthfulness in large lan

model-releasesarxiv-cs-ai
26 May 2026
Safety

TopoAlign: Topology-Aware Visual Representation Alignment

DGX agent

arXiv:2605.25541v1 Announce Type: cross Abstract: Neural networks encode inputs as high-dimensional vectors, known as representations, that capture how models process data by encoding task-relevant st

safetyarxiv-cs-ai
26 May 2026
Research

Towards Evaluation Engineering: An Empirical Study of ML Evaluation Harnesses in the Wild

DGX agent

arXiv:2605.24213v1 Announce Type: cross Abstract: Evaluation harnesses are software systems that orchestrate model evaluation by managing model invocation, data loading, metric computation, and result

researcharxiv-cs-ai
26 May 2026
Model Releases

Truthful Online Preference Aggregation for LLM Fine-Tuning in Mobile Crowdsourcing

DGX agent

arXiv:2605.24052v1 Announce Type: cross Abstract: To better serve users' demands in mobile applications (e.g., navigation), mobile crowdsourcing platforms can iteratively align large language model (L

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Understanding Conversational Patterns in Multi-agent Programming: A Case Study on Fibonacci Game Development

DGX agent

arXiv:2605.24138v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly applied to software engineering (SE), yet their potential for autonomous, role-oriented collaboration re

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

A Reproducible Universal Dependencies-Style Pipeline for Katharevousa Greek Parliamentary Text

DGX agent

arXiv:2605.22978v1 Announce Type: new Abstract: Katharevousa Greek remains poorly served by contemporary NLP pipelines despite its importance for legal, administrative, and parliamentary archives. We

model-releasesarxiv-cs-cl
25 May 2026
Research

Asymmetric Scaling Laws from Sparse Features

DGX agent

arXiv:2605.23591v1 Announce Type: cross Abstract: We introduce a model for neural scaling laws under sparse activations. In the model, test loss is often dominated by rare coordinates that are never o

researcharxiv-cs-lg
25 May 2026
Model Releases

Benchmarking and Enhancing VLM for Compressed Image Understanding

DGX agent

arXiv:2512.20901v2 Announce Type: replace Abstract: With the rapid development of Vision-Language Models (VLMs) and the growing demand for their applications, efficient compression of the image inputs

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Efficient and Transferable Agentic Knowledge Graph RAG via Reinforcement Learning

DGX agent

arXiv:2509.26383v5 Announce Type: replace-cross Abstract: Knowledge-graph retrieval-augmented generation (KG-RAG) couples large language models (LLMs) with structured, verifiable knowledge graphs (KGs

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation

DGX agent

arXiv:2510.08945v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) has emerged as a promising paradigm for improving factual accuracy in large language models (LLMs). We introduc

model-releasesarxiv-cs-ai
25 May 2026
Research

Is TabPFN the Silver Bullet for Insurance Pricing?

DGX agent

arXiv:2605.22892v1 Announce Type: cross Abstract: Modelling claim frequency and severity for non-life insurance pricing predominantly relies on generalised linear models, with gradient-boosted machine

researcharxiv-cs-lg
25 May 2026
Model Releases

Memorization Dynamics of Fill-in-the-Middle Pretraining

DGX agent

arXiv:2605.22981v1 Announce Type: cross Abstract: Fill-in-the-middle (FIM) is a pretraining objective widely used to equip causal language models with infilling ability, yet its effect on verbatim mem

model-releasesarxiv-cs-ai
25 May 2026
Research

Mitigating Object Hallucinations via Sentence-Level Early Intervention

DGX agent

arXiv:2507.12455v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have revolutionized cross-modal understanding but continue to struggle with hallucinations - fabricated con

researcharxiv-cs-cv
25 May 2026
Model Releases

ModeSwitch-LLM: A Lightweight Phase-Aware Controller for Cross-Mode LLM Inference on a Single GPU

DGX agent

arXiv:2605.23057v1 Announce Type: cross Abstract: ModeSwitch-LLM is a lightweight request-boundary controller for improving single-GPU large language model inference efficiency by routing each request

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Multilingual Steering by Design: Multilingual Sparse Autoencoders and Principled Layer Selection

DGX agent

arXiv:2605.23036v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) enable feature-level mechanistic interpretability and activation steering in large language models (LLMs), but SAE-based lang

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

SciHorizon-GENE: Benchmarking LLM for Life Sciences Inference from Gene Knowledge to Functional Understanding

DGX agent

arXiv:2601.12805v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown growing promise in biomedical research, particularly for knowledge-driven interpretation tasks. Howeve

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Sparse Autoencoders Map Brain-LLM Alignment onto Cortical Semantic Topography

DGX agent

arXiv:2605.23035v1 Announce Type: cross Abstract: Intermediate layers of large language models (LLMs) best predict human brain responses to language, one of the most robust findings in computational n

model-releasesarxiv-cs-ai
25 May 2026
Research

Valid and Expressive Copulas for Irregular Multivariate Time Series

DGX agent

arXiv:2605.23632v1 Announce Type: new Abstract: We introduce CopFITi, a copula model for probabilistic forecasting of irregular multivariate time series (IMTS). Our model combines the expressivity of

researcharxiv-cs-lg
25 May 2026
Model Releases

VisAnalog: A Diagnostic Suite for Visual Concept Transfer on Natural Images

DGX agent

arXiv:2605.23141v1 Announce Type: new Abstract: A useful test of visual concept learning is not just whether a model can recognize a concept in a single image, but whether it can preserve and manipula

model-releasesarxiv-cs-cv
25 May 2026
Agents

X-TRACK: Physics-Aware xLSTM for Realistic Vehicle Trajectory Prediction

DGX agent

arXiv:2511.00266v2 Announce Type: replace Abstract: Accurate trajectory prediction is crucial for safe and reliable autonomous driving systems, requiring models that capture long-term temporal depende

agentsarxiv-cs-lg
25 May 2026
Model Releases

A2QTGN: Adaptive Amplitude Quantum-Integrated Temporal Graph Network for Dynamic Link Prediction

DGX agent

arXiv:2605.21916v1 Announce Type: cross Abstract: Dynamic link prediction is important for modeling evolving interactions in complex systems, including social, communication, financial, and transporta

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

MapTab: Are MLLMs Ready for Multi-Criteria Route Planning in Heterogeneous Graphs?

DGX agent

arXiv:2602.18600v3 Announce Type: replace Abstract: Systematic evaluation of Multimodal Large Language Models (MLLMs) is crucial for advancing Artificial General Intelligence (AGI). However, existing

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Memory-Efficient LLM Pretraining via Minimalist Optimizer Design

DGX agent

arXiv:2506.16659v3 Announce Type: replace Abstract: Training large language models (LLMs) relies on adaptive optimizers such as Adam, which introduce extra operations and require significantly more me

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Rethinking Forward Processes for Score-Based Nonlinear Data Assimilation in High Dimensions

DGX agent

arXiv:2604.02889v2 Announce Type: replace-cross Abstract: Data assimilation is the process of estimating the state of a dynamical system over time by combining model predictions with measurements. Thi

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

SeqLoRA: Bilevel Orthogonal Adaptation for Continual Multi-Concept Generation

DGX agent

arXiv:2605.22743v1 Announce Type: new Abstract: Parameter-efficient fine-tuning enables fast personalization of text-to-image diffusion models, but composing multiple custom concepts remains challengi

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Why SGD is not Brownian Motion: A New Perspective on Stochastic Dynamics

DGX agent

arXiv:2605.22644v1 Announce Type: new Abstract: Stochastic Gradient Descent (SGD) is commonly modeled as a Langevin process, assuming that minibatch noise acts as Brownian motion. However, this approx

model-releasesarxiv-cs-lg
23 May 2026
Research

Bernini: Latent Semantic Planning for Video Diffusion

DGX agent

arXiv:2605.22344v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) and diffusion models have each reached remarkable maturity: MLLMs excel at reasoning over heterogeneous multimo

researcharxiv-cs-cv
22 May 2026
Model Releases

Beyond Benchmark Islands: Toward Representative Trustworthiness Evaluation for Agentic AI

DGX agent

arXiv:2603.14987v2 Announce Type: replace Abstract: Agentic AI systems increasingly act through tool-augmented, multi-step workflows whose failures (unsafe tool use, unauthorised actions, social harm)

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Blind Spots in the Guard: How Domain-Camouflaged Injection Attacks Evade Detection in Multi-Agent LLM Systems

DGX agent

arXiv:2605.22001v1 Announce Type: cross Abstract: Injection detectors deployed to protect LLM agents are calibrated on static, template-based payloads that announce themselves as override directives.

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Cross-Lingual Consensus: Aligning Multilingual Cultural Knowledge via Multilingual Self-Consistency

DGX agent

arXiv:2605.22137v1 Announce Type: new Abstract: Although Large Language Models (LLMs) demonstrate strong capabilities across various tasks, they exhibit significant performance discrepancies across la

model-releasesarxiv-cs-cl
22 May 2026
Research

From TF-IDF to Transformers: A Comparative and Ensemble Approach to Sentiment Classification

DGX agent

arXiv:2605.22003v1 Announce Type: new Abstract: Sentiment analysis, also referred to as opinion mining, primarily tries to extract opinion from any text-based data. In the context of movie reviews and

researcharxiv-cs-cl
22 May 2026
Model Releases

GHI: Graphormer over Conditioned Hypergraph Incidence for Aspect-Based Sentiment Analysis

DGX agent

arXiv:2605.22228v1 Announce Type: new Abstract: Aspect-based sentiment analysis (ABSA) requires models to bind sentiment evidence to the correct aspect, making it a natural testbed for fine-grained st

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

H-Flow: Self-supervised Human Scene Flow via Physics-inspired Joint Multi-modal Learning

DGX agent

arXiv:2605.22629v1 Announce Type: new Abstract: Parametric human models capture global pose but cannot represent the non-rigid surface dynamics of clothing and soft tissue. Generic scene flow estimate

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

InteractScience: Programmatic and Visually-Grounded Evaluation of Interactive Scientific Demonstration Code Generation

DGX agent

arXiv:2510.09724v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly capable of generating complete applications from natural language instructions, creating new opp

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning

DGX agent

arXiv:2605.21988v1 Announce Type: new Abstract: Video large language models (Video LLMs) achieve strong benchmark accuracy, yet often answer video questions through shortcuts such as single-frame cues

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

LongVT: Incentivizing 'Thinking with Long Videos' via Native Tool Calling

DGX agent

arXiv:2511.20785v3 Announce Type: replace Abstract: Large multimodal models (LMMs) have shown great potential for video reasoning with textual Chain-of-Thought. However, they remain vulnerable to hall

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

M3: Conversational LLMs Simplify Secure Clinical Data Access, Understanding, and Analysis

DGX agent

arXiv:2507.01053v4 Announce Type: replace-cross Abstract: Large-scale clinical databases offer opportunities for medical research, but their complexity creates barriers to effective use. The Medical I

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

MotiMotion: Motion-Controlled Video Generation with Visual Reasoning

DGX agent

arXiv:2605.22818v1 Announce Type: new Abstract: Current motion-controlled image-to-video generation models rigidly follow user-provided trajectories that are often sparse, imprecise, and causally inco

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality?

DGX agent

arXiv:2605.22109v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are increasingly deployed in human-facing roles where personality perception is critical, yet existing benchm

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

RAGCap-Bench: Benchmarking Capabilities of LLMs in Agentic Retrieval Augmented Generation Systems

DGX agent

arXiv:2510.13910v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) mitigates key limitations of Large Language Models (LLMs)-such as factual errors, outdated knowledge, and hallu

model-releasesarxiv-cs-cl
22 May 2026
← Previous
1…365366367368369…1074
Next →