AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,082 results
20 Apr 2026

UniEditBench: A Unified and Cost-Effective Benchmark for Image and Video Editing via Distilled MLLMs

Model ReleasesDGX agent

arXiv:2604.15871v1 Announce Type: cross Abstract: The evaluation of visual editing models remains fragmented across methods and modalities. Existing benchmarks are often tailored to specific paradigms

What Makes LLMs Effective Sequential Recommenders? A Study on Preference Intensity and Temporal Context

SafetyDGX agent

arXiv:2506.02261v3 Announce Type: replace-cross Abstract: What enables large language models (LLMs) to effectively model user preferences in sequential recommendation? Our investigation reveals that e

Why Fine-Tuning Encourages Hallucinations and How to Fix It

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.15574v1 Announce Type: cross Abstract: Large language models are prone to hallucinating factually incorrect statements. A key source of these errors is exposure to new factual information t

19 Apr 2026

You ever go on Huggingface and see: - GGUF - Unsloth - Llama.cpp - Dynamic GGUF - Q_4_M / IQ_4XL etc. Here's what's going on under the hood.…

Model ReleasesDGX agent

This post explains the technical details behind common terms and tools encountered on Hugging Face for running large language models locally, including quantization formats (GGUF, Q_4_M, IQ_4XL), opti

17 Apr 2026

Benchmarking Linguistic Adaptation in Comparable-Sized LLMs: A Study of Llama-3.1-8B, Mistral-7B-v0.1, and Qwen3-8B on Romanized Nepali

Model ReleasesDGX agent

arXiv:2604.14171v1 Announce Type: new Abstract: Romanized Nepali, the Nepali language written in the Latin alphabet, is the dominant medium for informal digital communication in Nepal, yet it remains

Between a Rock and a Hard Place: The Tension Between Ethical Reasoning and Safety Alignment in LLMs

SafetyDGX agent

arXiv:2509.05367v4 Announce Type: replace-cross Abstract: Large Language Model safety alignment predominantly operates on a binary assumption that requests are either safe or unsafe. This classificati

Beyond Translation: Evaluating Mathematical Reasoning Capabilities of LLMs in Sinhala and Tamil

ResearchDGX agent

arXiv:2602.14517v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong results in mathematical reasoning, and are increasingly deployed as tutoring and learning support

ConfLayers: Adaptive Confidence-based Layer Skipping for Self-Speculative Decoding

SafetyDGX agent

arXiv:2604.14612v1 Announce Type: cross Abstract: Self-speculative decoding is an inference technique for large language models designed to speed up generation without sacrificing output quality. It c

FoodSense: A Multisensory Food Dataset and Benchmark for Predicting Taste, Smell, Texture, and Sound from Images

Model ReleasesDGX agent

arXiv:2604.14388v1 Announce Type: new Abstract: Humans routinely infer taste, smell, texture, and even sound from food images a phenomenon well studied in cognitive science. However, prior vision lang

Formalizing the Safety, Security, and Functional Properties of Agentic AI Systems

SafetyDGX agent

arXiv:2510.14133v2 Announce Type: replace Abstract: Agentic AI systems, which leverage multiple autonomous agents and large language models (LLMs), are increasingly used to address complex, multi-step

From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning

ResearchDGX agent

arXiv:2604.15244v1 Announce Type: new Abstract: Speculative decoding (SD) accelerates large language model inference by allowing a lightweight draft model to propose outputs that a stronger target mod

Integrating Object Detection, LiDAR-Enhanced Depth Estimation, and Segmentation Models for Railway Environments

SafetyDGX agent

arXiv:2604.14781v1 Announce Type: new Abstract: Obstacle detection in railway environments is crucial for ensuring safety. However, very few studies address the problem using a complete, modular, and

Magnitude Is All You Need? Rethinking Phase in Quantum Encoding of Complex SAR Data

Model ReleasesDGX agent

arXiv:2604.14229v1 Announce Type: cross Abstract: Synthetic Aperture Radar (SAR) data is inherently complex-valued, while quantum machine learning (QML) models naturally operate in complex Hilbert spa

MARCA: A Checklist-Based Benchmark for Multilingual Web Search

Model ReleasesDGX agent

arXiv:2604.14448v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as sources of information, yet their reliability depends on the ability to search the web, select rel

Mechanistic Decoding of Cognitive Constructs in LLMs

Model ReleasesDGX agent

arXiv:2604.14593v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate increasingly sophisticated affective capabilities, the internal mechanisms by which they process complex

MIND: AI Co-Scientist for Material Research

AgentsDGX agent

arXiv:2604.13699v1 Announce Type: cross Abstract: Large language models (LLMs) have enabled agentic AI systems for scientific discovery, but most approaches remain limited to textbased reasoning witho

Physically-Induced Atmospheric Adversarial Perturbations: Enhancing Transferability and Robustness in Remote Sensing Image Classification

Model ReleasesDGX agent

arXiv:2604.14643v1 Announce Type: new Abstract: Adversarial attacks pose a severe threat to the reliability of deep learning models in remote sensing (RS) image classification. Most existing methods r

Shuffle the Context: RoPE-Perturbed Self-Distillation for Long-Context Adaptation

Model ReleasesDGX agent

arXiv:2604.14339v1 Announce Type: new Abstract: Large language models (LLMs) increasingly operate in settings that require reliable long-context understanding, such as retrieval-augmented generation a

SOLIS: Physics-Informed Learning of Interpretable Neural Surrogates for Nonlinear Systems

Model ReleasesDGX agent

arXiv:2604.14879v1 Announce Type: new Abstract: Nonlinear system identification must balance physical interpretability with model flexibility. Classical methods yield structured, control-relevant mode

TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation

ApplicationsDGX agent

arXiv:2604.14580v1 Announce Type: new Abstract: Existing audio-driven video digital human generation models rely on multi-step denoising, resulting in substantial computational overhead that severely

When Flat Minima Fail: Characterizing INT4 Quantization Collapse After FP32 Convergence

Model ReleasesDGX agent

arXiv:2604.15167v1 Announce Type: new Abstract: Post-training quantization (PTQ) assumes that a well-converged model is a quantization-ready model. We show this assumption fails in a structured, measu

16 Apr 2026

Enhancing Confidence Estimation in Telco LLMs via Twin-Pass CoT-Ensembling

Model ReleasesDGX agent

arXiv:2604.13271v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly applied to complex telecommunications tasks, including 3GPP specification analysis and O-RAN network troub

Free Geometry: Refining 3D Reconstruction from Longer Versions of Itself

Model ReleasesDGX agent

arXiv:2604.14048v1 Announce Type: new Abstract: Feed-forward 3D reconstruction models are efficient but rigid: once trained, they perform inference in a zero-shot manner and cannot adapt to the test s

From Weights to Activations: Is Steering the Next Frontier of Adaptation?

Model ReleasesDGX agent

arXiv:2604.14090v1 Announce Type: new Abstract: Post-training adaptation of language models is commonly achieved through parameter updates or input-based methods such as fine-tuning, parameter-efficie

Go from blank slate to analysis with BigQuery Studio notebook gallery templates

Model ReleasesDGX agent

For many data professionals, the most daunting part of a new project isn't the complexity of the data or the sophistication of the model, it’s the 'blank slate.' Staring at an empty notebook while cre

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Re…

Model ReleasesDGX agent

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Refusal rate: 98.8% down to 2.1% (!!) → Compliance: 1.2% up to

Identifiability of Potentially Degenerate Gaussian Mixture Models With Piecewise Affine Mixing

ResearchDGX agent

arXiv:2604.13218v1 Announce Type: cross Abstract: Causal representation learning (CRL) aims to identify the underlying latent variables from high-dimensional observations, even when variables are depe

Introducing GPT-Rosalind for life sciences research

Model ReleasesDGX agent

GPT-Rosalind is an AI model developed by OpenAI specifically designed to assist with life sciences research tasks. The model is trained to help researchers with applications such as analyzing biologic

Mathematical Reasoning Enhanced LLM for Formula Derivation: A Case Study on Fiber NLI Modellin

ApplicationsDGX agent

arXiv:2604.13062v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have demonstrated strong capabilities in code generation and text synthesis, yet their potential for sym

Mitigating Catastrophic Forgetting in Target Language Adaptation of LLMs via Source-Shielded Updates

Model ReleasesDGX agent

arXiv:2512.04844v2 Announce Type: replace Abstract: Expanding the linguistic diversity of instruct large language models (LLMs) is crucial for global accessibility but is often hindered by the relianc

Qwen 3.6 is here, and open-source! Run it locally with improved agentic coding capabilities. Try it with Claude Code: ollama launch claude -…

Model ReleasesDGX agent

Qwen 3.6 is here, and open-source! Run it locally with improved agentic coding capabilities. Try it with Claude Code: ollama launch claude --model qwen3.6 Try it with OpenClaw: ollama launch openclaw

RAG or Learning? Understanding the Limits of LLM Adaptation under Continuous Knowledge Drift in the Real World

Model ReleasesDGX agent

arXiv:2604.05096v2 Announce Type: replace Abstract: Large language models (LLMs) acquire most of their knowledge during pretraining, which ties them to a fixed snapshot of the world and makes adaptati

ReConText3D: Replay-based Continual Text-to-3D Generation

Model ReleasesDGX agent

arXiv:2604.13730v1 Announce Type: new Abstract: Continual learning enables models to acquire new knowledge over time while retaining previously learned capabilities. However, its application to text-t

Scalable Spatiotemporal Inference with Biased Scan Attention Transformer Neural Processes

HardwareDGX agent

arXiv:2506.09163v2 Announce Type: replace Abstract: Neural Processes (NPs) are a rapidly evolving class of models designed to directly model the posterior predictive distribution of stochastic process

Towards Successful Implementation of Automated Raveling Detection: Effects of Training Data Size, Illumination Difference, and Spatial Shift

Model ReleasesDGX agent

arXiv:2604.13322v1 Announce Type: new Abstract: Raveling, the loss of aggregates, is a major form of asphalt pavement surface distress, especially on highways. While research has shown that machine le

WAI-ANIMA 1.0 released

Model ReleasesDGX agent

WAI-ANIMA 1.0 is a newly released Stable Diffusion checkpoint model from the WAI model family, likely combining elements of the WAI-Illustrious anime generation lineage with the Anima diffusion archit

Why Multimodal In-Context Learning Lags Behind? Unveiling the Inner Mechanisms and Bottlenecks

SafetyDGX agent

arXiv:2604.13403v1 Announce Type: new Abstract: In-context learning (ICL) enables models to adapt to new tasks via inference-time demonstrations. Despite its success in large language models, the exte

ZK-APEX: Zero-Knowledge Approximate Personalized Unlearning with Executable Proofs

SafetyDGX agent

arXiv:2512.09953v2 Announce Type: replace-cross Abstract: Machine unlearning aims to remove the influence of specific data points from a trained model to satisfy privacy, copyright, and safety require

15 Apr 2026

A Dataset and Evaluation for Complex 4D Markerless Human Motion Capture

SafetyDGX agent

arXiv:2604.12765v1 Announce Type: new Abstract: Marker-based motion capture (MoCap) systems have long been the gold standard for accurate 4D human modeling, yet their reliance on specialized hardware

Coding-Free and Privacy-Preserving MCP Framework for Clinical Agentic Research Intelligence System

AgentsDGX agent

arXiv:2604.12258v1 Announce Type: cross Abstract: Clinical research involves labor-intensive processes such as study design, cohort construction, model development, and documentation, requiring domain

Cooperative Memory Paging with Keyword Bookmarks for Long-Horizon LLM Conversations

Model ReleasesDGX agent

arXiv:2604.12376v1 Announce Type: cross Abstract: When LLM conversations grow beyond the context window, old content must be evicted -- but how does the model recover it when needed? We propose cooper

DiffusionPrint: Learning Generative Fingerprints for Diffusion-Based Inpainting Localization

Local AiDGX agent

arXiv:2604.12443v1 Announce Type: new Abstract: Modern diffusion-based inpainting models pose significant challenges for image forgery localization (IFL), as their full regeneration pipelines reconstr

EDGE-Shield: Efficient Denoising-staGE Shield for Violative Content Filtering via Scalable Reference-Based Matching

Model ReleasesDGX agent

arXiv:2604.06063v2 Announce Type: replace Abstract: The advent of Text-to-Image generative models poses significant risks of copyright violation and deepfake generation. Since the rapid proliferation

EgoEsportsQA: An Egocentric Video Benchmark for Perception and Reasoning in Esports

Model ReleasesDGX agent

arXiv:2604.12320v1 Announce Type: cross Abstract: While video large language models (Video-LLMs) excel in understanding slow-paced, real-world egocentric videos, their capabilities in high-velocity, i

Evaluating Differential Privacy Against Membership Inference in Federated Learning: Insights from the NIST Genomics Red Team Challenge

Model ReleasesDGX agent

arXiv:2604.12737v1 Announce Type: cross Abstract: While Federated Learning (FL) mitigates direct data exposure, the resulting trained models remain susceptible to membership inference attacks (MIAs).

Fine-tuning Factor Augmented Neural Lasso for Heterogeneous Environments

Model ReleasesDGX agent

arXiv:2604.12288v1 Announce Type: cross Abstract: Fine-tuning is a widely used strategy for adapting pre-trained models to new tasks, yet its methodology and theoretical properties in high-dimensional

Generating Effective CoT Traces for Mitigating Causal Hallucination

TutorialsDGX agent

arXiv:2604.12748v1 Announce Type: new Abstract: Although large language models (LLMs) excel in complex reasoning tasks, they suffer from severe causal hallucination in event causality identification (

GUIDE: Guided Updates for In-context Decision Evolution in LLM-Driven Spacecraft Operations

SafetyDGX agent

arXiv:2603.27306v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been proposed as supervisory agents for spacecraft operations, but existing approaches rely on static prompt

iTeach: In the Wild Interactive Teaching for Failure-Driven Adaptation of Robot Perception

Model ReleasesDGX agent

arXiv:2410.09072v4 Announce Type: replace Abstract: Robotic perception models often fail when deployed in real-world environments due to out-of-distribution conditions such as clutter, occlusion, and

JanusCoder: Towards a Foundational Visual-Programmatic Interface for Code Intelligence

ResearchDGX agent

arXiv:2510.23538v2 Announce Type: replace Abstract: The scope of neural code intelligence is rapidly expanding beyond text-based source code to encompass the rich visual outputs that programs generate

KG-Hopper: Empowering Compact Open LLMs with Knowledge Graph Reasoning via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2603.21440v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) demonstrate impressive natural language capabilities but often struggle with knowledge-intensive reasoning tasks.

Masked by Consensus: Disentangling Privileged Knowledge in LLM Correctness

Local AiDGX agent

arXiv:2604.12373v1 Announce Type: new Abstract: Humans use introspection to evaluate their understanding through private internal states inaccessible to external observers. We investigate whether larg

Multi-region endpoints are available for Claude on Vertex AI

Model ReleasesDGX agent

Today, we’re announcing U.S. and EU multi-region endpoints for Claude on Vertex AI are available in public preview. By pooling capacity across multiple regions, these endpoints dynamically route reque

One Token Away from Collapse: The Fragility of Instruction-Tuned Helpfulness

ResearchDGX agent

arXiv:2604.13006v1 Announce Type: cross Abstract: Instruction-tuned large language models produce helpful, structured responses, but how robust is this helpfulness when trivially constrained? We show

QuarkMedSearch: A Long-Horizon Deep Search Agent for Exploring Medical Intelligence

Model ReleasesDGX agent

arXiv:2604.12867v1 Announce Type: new Abstract: As agentic foundation models continue to evolve, how to further improve their performance in vertical domains has become an important challenge. To this

SpecBound: Adaptive Bounded Self-Speculation with Layer-wise Confidence Calibration

ResearchDGX agent

arXiv:2604.12247v1 Announce Type: cross Abstract: Speculative decoding has emerged as a promising approach to accelerate autoregressive inference in large language models (LLMs). Self-draft methods, w

VISTA: Validation-Informed Trajectory Adaptation via Self-Distillation

ResearchDGX agent

arXiv:2604.12044v1 Announce Type: cross Abstract: Deep learning models may converge to suboptimal solutions despite strong validation accuracy, masking an optimization failure we term Trajectory Devia

Why and When Visual Token Pruning Fails? A Study on Relevant Visual Information Shift in MLLMs Decoding

ResearchDGX agent

arXiv:2604.12358v1 Announce Type: new Abstract: Recently, visual token pruning has been studied to handle the vast number of visual tokens in Multimodal Large Language Models. However, we observe that

14 Apr 2026

Adaptive Multi-Expert Reasoning via Difficulty-Aware Routing and Uncertainty-Guided Aggregation

ResearchDGX agent

arXiv:2604.10335v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong performance in math reasoning benchmarks, but their performance varies inconsistently across problems wi

AI Achieves a Perfect LSAT Score

ApplicationsDGX agent

arXiv:2604.10034v1 Announce Type: new Abstract: This paper reports the first documented instance of a language model achieving a perfect score on an officially disclosed Law School Admission Test (LSA

← Previous
1…290291292293294…1035
Next →