AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,221 results
Model Releases

SigGate-GT: Taming Over-Smoothing in Graph Transformers via Sigmoid-Gated Attention

DGX agent

arXiv:2604.17324v1 Announce Type: new Abstract: Graph transformers achieve strong results on molecular and long-range reasoning tasks, yet remain hampered by over-smoothing (the progressive collapse o

model-releasesarxiv-cs-lg
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Source-Free Domain Adaptation with Vision-Language Prior

DGX agent

arXiv:2604.17748v1 Announce Type: new Abstract: Source-Free Domain Adaptation (SFDA) seeks to adapt a source model, which is pre-trained on a supervised source domain, for a target domain, with only a

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

Spike-NVPT: Learning Robust Visual Prompts via Bio-Inspired Temporal Filtering and Discretization

DGX agent

arXiv:2604.18284v1 Announce Type: new Abstract: Pre-trained vision models have found widespread application across diverse domains. Prompt tuning-based methods have emerged as a parameter-efficient pa

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Stable On-Policy Distillation through Adaptive Target Reformulation

DGX agent

arXiv:2601.07155v2 Announce Type: replace Abstract: Knowledge distillation (KD) is a widely adopted technique for transferring knowledge from large language models to smaller student models; however,

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Style over Story: Measuring LLM Narrative Preferences via Structured Selection

DGX agent

arXiv:2510.02025v4 Announce Type: replace Abstract: We introduce a constraint-selection-based experiment design for measuring narrative preferences of Large Language Models (LLMs). This design offers

researcharxiv-cs-cl
21 Apr 2026
Research

Test-Time Reasoners Are Strategic Multiple-Choice Test-Takers

DGX agent

arXiv:2510.07761v2 Announce Type: replace Abstract: Large language models (LLMs) now give reasoning before answering, excelling in tasks like multiple-choice question answering (MCQA). Yet, a concern

researcharxiv-cs-cl
21 Apr 2026
Model Releases

The Geometric Canary: Predicting Steerability and Detecting Drift via Representational Stability

DGX agent

arXiv:2604.17698v1 Announce Type: cross Abstract: Reliable deployment of language models requires two capabilities that appear distinct but share a common geometric foundation: predicting whether a mo

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

Trustworthy deep domain adaptation for wearable photoplethysmography signal analysis with decision-theoretic uncertainty quantification

DGX agent

arXiv:2604.17480v1 Announce Type: new Abstract: In principle, deep generative models can be used to perform domain adaptation; i.e. align the input feature representations of test data with that of a

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

Uni-MMMU: A Massive Multi-discipline Multimodal Unified Benchmark

DGX agent

arXiv:2510.13759v3 Announce Type: replace Abstract: Unified multimodal models aim to jointly enable visual understanding and generation, yet current benchmarks rarely examine their true integration. E

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

1S-DAug: One-Shot Data Augmentation for Robust Few-Shot Generalization

DGX agent

arXiv:2602.00114v4 Announce Type: replace-cross Abstract: Few-shot learning (FSL) challenges model generalization to novel classes based on just a few shots of labeled examples, a testbed where tradit

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Context-Agent: Dynamic Discourse Trees for Non-Linear Dialogue

DGX agent

arXiv:2604.05552v2 Announce Type: replace-cross Abstract: Large Language Models demonstrate outstanding performance in many language tasks but still face fundamental challenges in managing the non-lin

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Facial-Expression-Aware Prompting for Empathetic LLM Tutoring

DGX agent

arXiv:2604.15336v1 Announce Type: cross Abstract: Large language models (LLMs) enable increasingly capable tutoring-style conversational agents, yet effective tutoring requires sensitivity to learners

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

From Benchmarking to Reasoning: A Dual-Aspect, Large-Scale Evaluation of LLMs on Vietnamese Legal Text

DGX agent

arXiv:2604.16270v1 Announce Type: cross Abstract: The complexity of Vietnam's legal texts presents a significant barrier to public access to justice. While Large Language Models offer a promising solu

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

FSPO: Few-Shot Optimization of Synthetic Preferences Personalizes to Real Users

DGX agent

arXiv:2502.19312v2 Announce Type: replace-cross Abstract: Effective personalization of LLMs is critical for a broad range of user-interfacing applications such as virtual assistants and content curati

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

Histogram-based Parameter-efficient Tuning for Passive and Active Sonar Classification

DGX agent

arXiv:2504.15214v3 Announce Type: replace Abstract: Parameter-efficient transfer learning (PETL) methods adapt large artificial neural networks to downstream tasks without fine-tuning the entire model

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

Interpretable Traces, Unexpected Outcomes: Investigating the Disconnect in Trace-Based Knowledge Distillation

DGX agent

arXiv:2505.13792v2 Announce Type: replace-cross Abstract: Recent advances in reasoning-focused Large Language Models (LLMs) have introduced Chain-of-Thought (CoT) traces - intermediate reasoning steps

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Kimi K2.6 now in OpenCode — Go included

DGX agent

Kimi K2.6, an AI model developed by Moonshot, has been released on OpenCode with support for the Go programming language. This release expands the model's capabilities to handle Go code generation, an

model-releaseskimi-moonshot--x
20 Apr 2026
Model Releases

LLMs Corrupt Your Documents When You Delegate

DGX agent

arXiv:2604.15597v1 Announce Type: new Abstract: Large Language Models (LLMs) are poised to disrupt knowledge work, with the emergence of delegated work as a new interaction paradigm (e.g., vibe coding

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Making Image Editing Easier via Adaptive Task Reformulation with Agentic Executions

DGX agent

arXiv:2604.15917v1 Announce Type: new Abstract: Instruction guided image editing has advanced substantially with recent generative models, yet it still fails to produce reliable results across many se

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Qwen3.5-Omni Technical Report

DGX agent

arXiv:2604.15804v1 Announce Type: new Abstract: In this work, we present Qwen3.5-Omni, the latest advancement in the Qwen-Omni model family. Representing a significant evolution over its predecessor,

model-releasesarxiv-cs-cl
20 Apr 2026
Safety

Self-Distillation as a Performance Recovery Mechanism for LLMs: Counteracting Compression and Catastrophic Forgetting

DGX agent

arXiv:2604.15794v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable success, underpinning diverse AI applications. However, they often suffer from performance degra

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

Stochasticity in Tokenisation Improves Robustness

DGX agent

arXiv:2604.16037v1 Announce Type: new Abstract: The widespread adoption of large language models (LLMs) has increased concerns about their robustness. Vulnerabilities in perturbations of tokenisation

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Time to really give Kimi K2.6 a go. Thank you @ollama! Love the ollama cloud setup!

DGX agent

Time to really give Kimi K2.6 a go. Thank you @ollama! Love the ollama cloud setup! Kimi K2.6 raises the bar for open-source models. 🦙 available on Ollama's cloud! Try it with OpenClaw: ollama launch

model-releasesollama--x
20 Apr 2026
Model Releases

Transformer Neural Processes - Kernel Regression

DGX agent

arXiv:2411.12502v4 Announce Type: replace-cross Abstract: Neural Processes (NPs) are a rapidly evolving class of models designed to directly model the posterior predictive distribution of stochastic p

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

UniEditBench: A Unified and Cost-Effective Benchmark for Image and Video Editing via Distilled MLLMs

DGX agent

arXiv:2604.15871v1 Announce Type: cross Abstract: The evaluation of visual editing models remains fragmented across methods and modalities. Existing benchmarks are often tailored to specific paradigms

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

What Makes LLMs Effective Sequential Recommenders? A Study on Preference Intensity and Temporal Context

DGX agent

arXiv:2506.02261v3 Announce Type: replace-cross Abstract: What enables large language models (LLMs) to effectively model user preferences in sequential recommendation? Our investigation reveals that e

safetyarxiv-cs-lg
20 Apr 2026
Model Releases

Why Fine-Tuning Encourages Hallucinations and How to Fix It

DGX agent

arXiv:2604.15574v1 Announce Type: cross Abstract: Large language models are prone to hallucinating factually incorrect statements. A key source of these errors is exposure to new factual information t

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

You ever go on Huggingface and see: - GGUF - Unsloth - Llama.cpp - Dynamic GGUF - Q_4_M / IQ_4XL etc. Here's what's going on under the hood.…

DGX agent

This post explains the technical details behind common terms and tools encountered on Hugging Face for running large language models locally, including quantization formats (GGUF, Q_4_M, IQ_4XL), opti

model-releasesclem-delangue--x
19 Apr 2026
Model Releases

Benchmarking Linguistic Adaptation in Comparable-Sized LLMs: A Study of Llama-3.1-8B, Mistral-7B-v0.1, and Qwen3-8B on Romanized Nepali

DGX agent

arXiv:2604.14171v1 Announce Type: new Abstract: Romanized Nepali, the Nepali language written in the Latin alphabet, is the dominant medium for informal digital communication in Nepal, yet it remains

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Between a Rock and a Hard Place: The Tension Between Ethical Reasoning and Safety Alignment in LLMs

DGX agent

arXiv:2509.05367v4 Announce Type: replace-cross Abstract: Large Language Model safety alignment predominantly operates on a binary assumption that requests are either safe or unsafe. This classificati

safetyarxiv-cs-ai
17 Apr 2026
Research

Beyond Translation: Evaluating Mathematical Reasoning Capabilities of LLMs in Sinhala and Tamil

DGX agent

arXiv:2602.14517v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong results in mathematical reasoning, and are increasingly deployed as tutoring and learning support

researcharxiv-cs-cl
17 Apr 2026
Safety

ConfLayers: Adaptive Confidence-based Layer Skipping for Self-Speculative Decoding

DGX agent

arXiv:2604.14612v1 Announce Type: cross Abstract: Self-speculative decoding is an inference technique for large language models designed to speed up generation without sacrificing output quality. It c

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

FoodSense: A Multisensory Food Dataset and Benchmark for Predicting Taste, Smell, Texture, and Sound from Images

DGX agent

arXiv:2604.14388v1 Announce Type: new Abstract: Humans routinely infer taste, smell, texture, and even sound from food images a phenomenon well studied in cognitive science. However, prior vision lang

model-releasesarxiv-cs-cv
17 Apr 2026
Safety

Formalizing the Safety, Security, and Functional Properties of Agentic AI Systems

DGX agent

arXiv:2510.14133v2 Announce Type: replace Abstract: Agentic AI systems, which leverage multiple autonomous agents and large language models (LLMs), are increasingly used to address complex, multi-step

safetyarxiv-cs-ai
17 Apr 2026
Research

From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning

DGX agent

arXiv:2604.15244v1 Announce Type: new Abstract: Speculative decoding (SD) accelerates large language model inference by allowing a lightweight draft model to propose outputs that a stronger target mod

researcharxiv-cs-cl
17 Apr 2026
Safety

Integrating Object Detection, LiDAR-Enhanced Depth Estimation, and Segmentation Models for Railway Environments

DGX agent

arXiv:2604.14781v1 Announce Type: new Abstract: Obstacle detection in railway environments is crucial for ensuring safety. However, very few studies address the problem using a complete, modular, and

safetyarxiv-cs-cv
17 Apr 2026
Model Releases

Magnitude Is All You Need? Rethinking Phase in Quantum Encoding of Complex SAR Data

DGX agent

arXiv:2604.14229v1 Announce Type: cross Abstract: Synthetic Aperture Radar (SAR) data is inherently complex-valued, while quantum machine learning (QML) models naturally operate in complex Hilbert spa

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

MARCA: A Checklist-Based Benchmark for Multilingual Web Search

DGX agent

arXiv:2604.14448v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as sources of information, yet their reliability depends on the ability to search the web, select rel

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Mechanistic Decoding of Cognitive Constructs in LLMs

DGX agent

arXiv:2604.14593v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate increasingly sophisticated affective capabilities, the internal mechanisms by which they process complex

model-releasesarxiv-cs-cl
17 Apr 2026
Agents

MIND: AI Co-Scientist for Material Research

DGX agent

arXiv:2604.13699v1 Announce Type: cross Abstract: Large language models (LLMs) have enabled agentic AI systems for scientific discovery, but most approaches remain limited to textbased reasoning witho

agentsarxiv-cs-ai
17 Apr 2026
Model Releases

Physically-Induced Atmospheric Adversarial Perturbations: Enhancing Transferability and Robustness in Remote Sensing Image Classification

DGX agent

arXiv:2604.14643v1 Announce Type: new Abstract: Adversarial attacks pose a severe threat to the reliability of deep learning models in remote sensing (RS) image classification. Most existing methods r

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Shuffle the Context: RoPE-Perturbed Self-Distillation for Long-Context Adaptation

DGX agent

arXiv:2604.14339v1 Announce Type: new Abstract: Large language models (LLMs) increasingly operate in settings that require reliable long-context understanding, such as retrieval-augmented generation a

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

SOLIS: Physics-Informed Learning of Interpretable Neural Surrogates for Nonlinear Systems

DGX agent

arXiv:2604.14879v1 Announce Type: new Abstract: Nonlinear system identification must balance physical interpretability with model flexibility. Classical methods yield structured, control-relevant mode

model-releasesarxiv-cs-lg
17 Apr 2026
Applications

TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation

DGX agent

arXiv:2604.14580v1 Announce Type: new Abstract: Existing audio-driven video digital human generation models rely on multi-step denoising, resulting in substantial computational overhead that severely

applicationsarxiv-cs-cv
17 Apr 2026
Model Releases

When Flat Minima Fail: Characterizing INT4 Quantization Collapse After FP32 Convergence

DGX agent

arXiv:2604.15167v1 Announce Type: new Abstract: Post-training quantization (PTQ) assumes that a well-converged model is a quantization-ready model. We show this assumption fails in a structured, measu

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

Enhancing Confidence Estimation in Telco LLMs via Twin-Pass CoT-Ensembling

DGX agent

arXiv:2604.13271v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly applied to complex telecommunications tasks, including 3GPP specification analysis and O-RAN network troub

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Free Geometry: Refining 3D Reconstruction from Longer Versions of Itself

DGX agent

arXiv:2604.14048v1 Announce Type: new Abstract: Feed-forward 3D reconstruction models are efficient but rigid: once trained, they perform inference in a zero-shot manner and cannot adapt to the test s

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

From Weights to Activations: Is Steering the Next Frontier of Adaptation?

DGX agent

arXiv:2604.14090v1 Announce Type: new Abstract: Post-training adaptation of language models is commonly achieved through parameter updates or input-based methods such as fine-tuning, parameter-efficie

model-releasesarxiv-cs-cl
16 Apr 2026
← Previous
1…376377378379380…1338
Next →