AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlog
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
Model Releases

RAG or Learning? Understanding the Limits of LLM Adaptation under Continuous Knowledge Drift in the Real World

DGX agent

arXiv:2604.05096v2 Announce Type: replace Abstract: Large language models (LLMs) acquire most of their knowledge during pretraining, which ties them to a fixed snapshot of the world and makes adaptati

model-releasesarxiv-cs-cl
16 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

ReConText3D: Replay-based Continual Text-to-3D Generation

DGX agent

arXiv:2604.13730v1 Announce Type: new Abstract: Continual learning enables models to acquire new knowledge over time while retaining previously learned capabilities. However, its application to text-t

model-releasesarxiv-cs-cv
16 Apr 2026
Hardware

Scalable Spatiotemporal Inference with Biased Scan Attention Transformer Neural Processes

DGX agent

arXiv:2506.09163v2 Announce Type: replace Abstract: Neural Processes (NPs) are a rapidly evolving class of models designed to directly model the posterior predictive distribution of stochastic process

hardwarearxiv-cs-lg
16 Apr 2026
Model Releases

Towards Successful Implementation of Automated Raveling Detection: Effects of Training Data Size, Illumination Difference, and Spatial Shift

DGX agent

arXiv:2604.13322v1 Announce Type: new Abstract: Raveling, the loss of aggregates, is a major form of asphalt pavement surface distress, especially on highways. While research has shown that machine le

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

WAI-ANIMA 1.0 released

DGX agent

WAI-ANIMA 1.0 is a newly released Stable Diffusion checkpoint model from the WAI model family, likely combining elements of the WAI-Illustrious anime generation lineage with the Anima diffusion archit

model-releasesr-stablediffusion
16 Apr 2026
Safety

Why Multimodal In-Context Learning Lags Behind? Unveiling the Inner Mechanisms and Bottlenecks

DGX agent

arXiv:2604.13403v1 Announce Type: new Abstract: In-context learning (ICL) enables models to adapt to new tasks via inference-time demonstrations. Despite its success in large language models, the exte

safetyarxiv-cs-cv
16 Apr 2026
Safety

ZK-APEX: Zero-Knowledge Approximate Personalized Unlearning with Executable Proofs

DGX agent

arXiv:2512.09953v2 Announce Type: replace-cross Abstract: Machine unlearning aims to remove the influence of specific data points from a trained model to satisfy privacy, copyright, and safety require

safetyarxiv-cs-lg
16 Apr 2026
Safety

A Dataset and Evaluation for Complex 4D Markerless Human Motion Capture

DGX agent

arXiv:2604.12765v1 Announce Type: new Abstract: Marker-based motion capture (MoCap) systems have long been the gold standard for accurate 4D human modeling, yet their reliance on specialized hardware

safetyarxiv-cs-cv
15 Apr 2026
Agents

Coding-Free and Privacy-Preserving MCP Framework for Clinical Agentic Research Intelligence System

DGX agent

arXiv:2604.12258v1 Announce Type: cross Abstract: Clinical research involves labor-intensive processes such as study design, cohort construction, model development, and documentation, requiring domain

agentsarxiv-cs-ai
15 Apr 2026
Model Releases

Cooperative Memory Paging with Keyword Bookmarks for Long-Horizon LLM Conversations

DGX agent

arXiv:2604.12376v1 Announce Type: cross Abstract: When LLM conversations grow beyond the context window, old content must be evicted -- but how does the model recover it when needed? We propose cooper

model-releasesarxiv-cs-ai
15 Apr 2026
Local Ai

DiffusionPrint: Learning Generative Fingerprints for Diffusion-Based Inpainting Localization

DGX agent

arXiv:2604.12443v1 Announce Type: new Abstract: Modern diffusion-based inpainting models pose significant challenges for image forgery localization (IFL), as their full regeneration pipelines reconstr

local-aiarxiv-cs-cv
15 Apr 2026
Model Releases

EDGE-Shield: Efficient Denoising-staGE Shield for Violative Content Filtering via Scalable Reference-Based Matching

DGX agent

arXiv:2604.06063v2 Announce Type: replace Abstract: The advent of Text-to-Image generative models poses significant risks of copyright violation and deepfake generation. Since the rapid proliferation

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

EgoEsportsQA: An Egocentric Video Benchmark for Perception and Reasoning in Esports

DGX agent

arXiv:2604.12320v1 Announce Type: cross Abstract: While video large language models (Video-LLMs) excel in understanding slow-paced, real-world egocentric videos, their capabilities in high-velocity, i

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Evaluating Differential Privacy Against Membership Inference in Federated Learning: Insights from the NIST Genomics Red Team Challenge

DGX agent

arXiv:2604.12737v1 Announce Type: cross Abstract: While Federated Learning (FL) mitigates direct data exposure, the resulting trained models remain susceptible to membership inference attacks (MIAs).

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Fine-tuning Factor Augmented Neural Lasso for Heterogeneous Environments

DGX agent

arXiv:2604.12288v1 Announce Type: cross Abstract: Fine-tuning is a widely used strategy for adapting pre-trained models to new tasks, yet its methodology and theoretical properties in high-dimensional

model-releasesarxiv-cs-lg
15 Apr 2026
Tutorials

Generating Effective CoT Traces for Mitigating Causal Hallucination

DGX agent

arXiv:2604.12748v1 Announce Type: new Abstract: Although large language models (LLMs) excel in complex reasoning tasks, they suffer from severe causal hallucination in event causality identification (

tutorialsarxiv-cs-cl
15 Apr 2026
Safety

GUIDE: Guided Updates for In-context Decision Evolution in LLM-Driven Spacecraft Operations

DGX agent

arXiv:2603.27306v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been proposed as supervisory agents for spacecraft operations, but existing approaches rely on static prompt

safetyarxiv-cs-ai
15 Apr 2026
Model Releases

iTeach: In the Wild Interactive Teaching for Failure-Driven Adaptation of Robot Perception

DGX agent

arXiv:2410.09072v4 Announce Type: replace Abstract: Robotic perception models often fail when deployed in real-world environments due to out-of-distribution conditions such as clutter, occlusion, and

model-releasesarxiv-cs-ro
15 Apr 2026
Research

JanusCoder: Towards a Foundational Visual-Programmatic Interface for Code Intelligence

DGX agent

arXiv:2510.23538v2 Announce Type: replace Abstract: The scope of neural code intelligence is rapidly expanding beyond text-based source code to encompass the rich visual outputs that programs generate

researcharxiv-cs-ai
15 Apr 2026
Model Releases

KG-Hopper: Empowering Compact Open LLMs with Knowledge Graph Reasoning via Reinforcement Learning

DGX agent

arXiv:2603.21440v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) demonstrate impressive natural language capabilities but often struggle with knowledge-intensive reasoning tasks.

model-releasesarxiv-cs-ai
15 Apr 2026
Local Ai

Masked by Consensus: Disentangling Privileged Knowledge in LLM Correctness

DGX agent

arXiv:2604.12373v1 Announce Type: new Abstract: Humans use introspection to evaluate their understanding through private internal states inaccessible to external observers. We investigate whether larg

local-aiarxiv-cs-cl
15 Apr 2026
Model Releases

Multi-region endpoints are available for Claude on Vertex AI

DGX agent

Today, we’re announcing U.S. and EU multi-region endpoints for Claude on Vertex AI are available in public preview. By pooling capacity across multiple regions, these endpoints dynamically route reque

model-releasesgoogle-cloud-ai
15 Apr 2026
Research

One Token Away from Collapse: The Fragility of Instruction-Tuned Helpfulness

DGX agent

arXiv:2604.13006v1 Announce Type: cross Abstract: Instruction-tuned large language models produce helpful, structured responses, but how robust is this helpfulness when trivially constrained? We show

researcharxiv-cs-ai
15 Apr 2026
Model Releases

QuarkMedSearch: A Long-Horizon Deep Search Agent for Exploring Medical Intelligence

DGX agent

arXiv:2604.12867v1 Announce Type: new Abstract: As agentic foundation models continue to evolve, how to further improve their performance in vertical domains has become an important challenge. To this

model-releasesarxiv-cs-ai
15 Apr 2026
Research

SpecBound: Adaptive Bounded Self-Speculation with Layer-wise Confidence Calibration

DGX agent

arXiv:2604.12247v1 Announce Type: cross Abstract: Speculative decoding has emerged as a promising approach to accelerate autoregressive inference in large language models (LLMs). Self-draft methods, w

researcharxiv-cs-ai
15 Apr 2026
Research

VISTA: Validation-Informed Trajectory Adaptation via Self-Distillation

DGX agent

arXiv:2604.12044v1 Announce Type: cross Abstract: Deep learning models may converge to suboptimal solutions despite strong validation accuracy, masking an optimization failure we term Trajectory Devia

researcharxiv-cs-ai
15 Apr 2026
Research

Why and When Visual Token Pruning Fails? A Study on Relevant Visual Information Shift in MLLMs Decoding

DGX agent

arXiv:2604.12358v1 Announce Type: new Abstract: Recently, visual token pruning has been studied to handle the vast number of visual tokens in Multimodal Large Language Models. However, we observe that

researcharxiv-cs-cv
15 Apr 2026
Research

Adaptive Multi-Expert Reasoning via Difficulty-Aware Routing and Uncertainty-Guided Aggregation

DGX agent

arXiv:2604.10335v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong performance in math reasoning benchmarks, but their performance varies inconsistently across problems wi

researcharxiv-cs-cl
14 Apr 2026
Applications

AI Achieves a Perfect LSAT Score

DGX agent

arXiv:2604.10034v1 Announce Type: new Abstract: This paper reports the first documented instance of a language model achieving a perfect score on an officially disclosed Law School Admission Test (LSA

applicationsarxiv-cs-ai
14 Apr 2026
Research

Attention-Guided Flow-Matching for Sparse 3D Geological Generation

DGX agent

arXiv:2604.09700v1 Announce Type: cross Abstract: Constructing high-resolution 3D geological models from sparse 1D borehole and 2D surface data is a highly ill-posed inverse problem. Traditional heuri

researcharxiv-cs-ai
14 Apr 2026
Model Releases

BiCLIP: Domain Canonicalization via Structured Geometric Transformation

DGX agent

arXiv:2603.08942v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have demonstrated remarkable zero-shot capabilities, yet adapting these models to specialized

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CARINOX: Inference-time Scaling with Category-Aware Reward-based Initial Noise Optimization and Exploration

DGX agent

arXiv:2509.17458v3 Announce Type: replace-cross Abstract: Text-to-image diffusion models, such as Stable Diffusion, can produce high-quality and diverse images but often fail to achieve compositional

model-releasesarxiv-cs-cl
14 Apr 2026
Local Ai

Codex CLI with ollama as provider?

DGX agent

This r/ollama thread discusses using Ollama as a local model provider for OpenAI's Codex CLI, covering configuration and setup. Open models can be used with OpenAI's Codex CLI through Ollama, allowing

local-air-ollama
14 Apr 2026
Applications

CraftGraffiti: Exploring Human Identity with Custom Graffiti Art via Facial-Preserving Diffusion Models

DGX agent

arXiv:2508.20640v2 Announce Type: replace Abstract: Preserving facial identity under extreme stylistic transformation remains a major challenge in generative art. In graffiti, a high-contrast, abstrac

applicationsarxiv-cs-cv
14 Apr 2026
Model Releases

Data-Efficient Semantic Segmentation of 3D Point Clouds via Open-Vocabulary Image Segmentation-based Pseudo-Labeling

DGX agent

arXiv:2604.11007v1 Announce Type: new Abstract: Semantic segmentation of 3D point cloud scenes is a crucial task for various applications. In real-world scenarios, training segmentation models often f

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Decomposing and Reducing Hidden Measurement Error in LLM Evaluation Pipelines

DGX agent

arXiv:2604.11581v1 Announce Type: new Abstract: LLM evaluations drive which models get deployed, which safety standards get adopted, and which research conclusions get published. Yet these scores carr

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Disambiguation-Centric Finetuning Makes Enterprise Tool-Calling LLMs More Realistic and Less Risky

DGX agent

arXiv:2507.03336v4 Announce Type: replace Abstract: Large language models (LLMs) are increasingly tasked with invoking enterprise APIs, yet they routinely falter when near-duplicate tools vie for the

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Do Machines Fail Like Humans? A Human-Centred Out-of-Distribution Spectrum for Mapping Error Alignment

DGX agent

arXiv:2603.07462v2 Announce Type: replace Abstract: Determining whether AI systems process information similarly to humans is central to cognitive science and trustworthy AI. While modern AI models ca

safetyarxiv-cs-ai
14 Apr 2026
Research

Exploring Knowledge Purification in Multi-Teacher Knowledge Distillation for LLMs

DGX agent

arXiv:2602.01064v2 Announce Type: replace Abstract: Knowledge distillation has emerged as a pivotal technique for transferring knowledge from stronger large language models (LLMs) to smaller, more eff

researcharxiv-cs-cl
14 Apr 2026
Model Releases

FashionMV: Product-Level Composed Image Retrieval with Multi-View Fashion Data

DGX agent

arXiv:2604.10297v1 Announce Type: cross Abstract: Composed Image Retrieval (CIR) retrieves target images using a reference image paired with modification text. Despite rapid advances, all existing met

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Fatigue-PINN: Physics-Informed Fatigue-Driven Motion Modulation and Synthesis

DGX agent

arXiv:2502.19056v2 Announce Type: replace-cross Abstract: Fatigue modeling is essential for motion synthesis tasks to model human motions under fatigued conditions and biomechanical engineering applic

safetyarxiv-cs-lg
14 Apr 2026
Model Releases

FedRio: Personalized Federated Social Bot Detection via Cooperative Reinforced Contrastive Adversarial Distillation

DGX agent

arXiv:2604.10678v1 Announce Type: new Abstract: Social bot detection is critical to the stability and security of online social platforms. However, current state-of-the-art bot detection models are la

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

From GPT-3 to GPT-5: Mapping their capabilities, scope, limitations, and consequences

DGX agent

arXiv:2604.10332v1 Announce Type: new Abstract: We present the progress of the GPT family from GPT-3 through GPT-3.5, GPT-4, GPT-4 Turbo, GPT-4o, GPT-4.1, and the GPT-5 family. Our work is comparative

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Frugal Knowledge Graph Construction with Local LLMs: A Zero-Shot Pipeline, Self-Consistency and Wisdom of Artificial Crowds

DGX agent

arXiv:2604.11104v1 Announce Type: new Abstract: This paper presents an empirical study of a multi-model zero-shot pipeline for knowledge graph construction and exploitation, executed entirely through

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

GLEaN: A Text-to-image Bias Detection Approach for Public Comprehension

DGX agent

arXiv:2604.09923v1 Announce Type: new Abstract: Text-to-image (T2I) models, and their encoded biases, increasingly shape the visual media the public encounters. While researchers have produced a rich

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

[Help] Gemma 4 26B LoRA Training on 16GB VRAM: Loss decreases, but inference degenerates into loops (Masking vs. MoE?)

DGX agent

This Reddit thread discusses a user's experience attempting LoRA fine-tuning of the Gemma 4 26B-A4B model on a 16GB VRAM GPU, where training loss decreases normally but the resulting model degenerates

model-releasesr-ollama
14 Apr 2026
Model Releases

Learning Robustness at Test-Time from a Non-Robust Teacher

DGX agent

arXiv:2604.11590v1 Announce Type: new Abstract: Nowadays, pretrained models are increasingly used as general-purpose backbones and adapted at test-time to downstream environments where target data are

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LIFT: A Novel Framework for Enhancing Long-Context Understanding of LLMs via Long Input Fine-Tuning

DGX agent

arXiv:2502.14644v5 Announce Type: replace Abstract: Long context understanding remains challenging for large language models due to their limited context windows. This paper introduces Long Input Fine

model-releasesarxiv-cs-cl
14 Apr 2026
← Previous
1…373374375376377…1324
Next →