AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,098 results
Model Releases

Neural networks for Text-to-Speech evaluation

DGX agent

arXiv:2604.08562v1 Announce Type: cross Abstract: Ensuring that Text-to-Speech (TTS) systems deliver human-perceived quality at scale is a central challenge for modern speech technologies. Human subje

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Nexus: Same Pretraining Loss, Better Downstream Generalization via Common Minima

DGX agent

arXiv:2604.09258v1 Announce Type: new Abstract: Pretraining is the cornerstone of Large Language Models (LLMs), dominating the vast majority of computational budget and data to serve as the primary en

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Noise-Aware In-Context Learning for Hallucination Mitigation in ALLMs

DGX agent

arXiv:2604.09021v1 Announce Type: cross Abstract: Auditory large language models (ALLMs) have demonstrated strong general capabilities in audio understanding and reasoning tasks. However, their reliab

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: Multi-Exposure Image Fusion in Dynamic Scenes (Track 2)

DGX agent

arXiv:2604.09030v1 Announce Type: new Abstract: This paper presents NTIRE 2026, the 3rd Restore Any Image Model (RAIM) challenge on multi-exposure image fusion in dynamic scenes. We introduce a benchm

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Ocr benchmark

DGX agent

Ocr benchmark We’re open sourcing the first document OCR benchmark for the agentic era, ParseBench. Document parsing is the foundation of every AI agent that works with real-world files. ParseBench is

model-releasesjerry-liu--x
13 Apr 2026
Model Releases

Ollama 0.20.6 is here with improved Gemma 4 tool calling! more improvements to come for Gemma 4!

DGX agent

Ollama version 0.20.6 has been released, featuring improved tool calling support for Google's Gemma 4 model. The update focuses on enhancing the reliability and functionality of function/tool calling

model-releasesollama--x
13 Apr 2026
Model Releases

Ollama / Mistral with MCP to Mempalace

DGX agent

This Reddit post from r/ollama discusses integrating Ollama-served Mistral with MemPalace — a free, locally-run AI memory system — via the Model Context Protocol (MCP). MemPalace runs entirely on a us

model-releasesr-ollama
13 Apr 2026
Model Releases

On Semiotic-Grounded Interpretive Evaluation of Generative Art

DGX agent

arXiv:2604.08641v1 Announce Type: cross Abstract: Interpretation is essential to deciphering the language of art: audiences communicate with artists by recovering meaning from visual artifacts. Howeve

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

On-the-Fly Adaptation to Quantization: Configuration-Aware LoRA for Efficient Fine-Tuning of Quantized LLMs

DGX agent

arXiv:2509.25214v3 Announce Type: replace-cross Abstract: As increasingly large pre-trained models are released, deploying them on edge devices for privacy-preserving applications requires effective c

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Online Intention Prediction via Control-Informed Learning

DGX agent

arXiv:2604.09303v1 Announce Type: cross Abstract: This paper presents an online intention prediction framework for estimating the goal state of autonomous systems in real time, even when intention is

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Open-source AI is now matching GPT-4 — and you can run it privately for free

DGX agent

This Reddit post discusses how open-source AI models have advanced to rival GPT-4-level performance and can be run locally for free using Ollama — a tool that lets users download and manage large lang

model-releasesr-ollama
13 Apr 2026
Model Releases

PACED: Distillation and On-Policy Self-Distillation at the Frontier of Student Competence

DGX agent

arXiv:2603.11178v3 Announce Type: replace Abstract: Standard LLM distillation treats all training problems equally -- wasting compute on problems the student has already mastered or cannot yet solve.

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Parameterized Complexity Of Representing Models Of MSO Formulas

DGX agent

arXiv:2604.08707v1 Announce Type: new Abstract: Monadic second order logic (MSO2) plays an important role in parameterized complexity due to the Courcelle's theorem. This theorem states that the probl

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

ParseBench is here!📊 We’ve just released ParseBench, an open benchmark + dataset for evaluating document parsing at scale. It includes: • 2…

DGX agent

ParseBench is here!📊 We’ve just released ParseBench, an open benchmark + dataset for evaluating document parsing at scale. It includes: • 2,000+ human-reviewed enterprise documents • 167,000 evaluatio

model-releasesjerry-liu--x
13 Apr 2026
Model Releases

PhysInOne: Visual Physics Learning and Reasoning in One Suite

DGX agent

arXiv:2604.09415v1 Announce Type: cross Abstract: We present PhysInOne, a large-scale synthetic dataset addressing the critical scarcity of physically-grounded training data for AI systems. Unlike exi

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

PilotBench: A Benchmark for General Aviation Agents with Safety Constraints

DGX agent

arXiv:2604.08987v1 Announce Type: new Abstract: As Large Language Models (LLMs) advance toward embodied AI agents operating in physical environments, a fundamental question emerges: can models trained

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos

DGX agent

arXiv:2604.08991v1 Announce Type: cross Abstract: Small object-centric spatial understanding in indoor videos remains a significant challenge for multimodal large language models (MLLMs), despite its

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Precise Shield: Explaining and Aligning VLLM Safety via Neuron-Level Guidance

DGX agent

arXiv:2604.08881v1 Announce Type: new Abstract: In real-world deployments, Vision-Language Large Models (VLLMs) face critical challenges from multilingual and multimodal composite attacks: harmful ima

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Precomputing Multi-Agent Path Replanning using Temporal Flexibility

DGX agent

arXiv:2601.04884v2 Announce Type: replace Abstract: Executing a multi-agent plan can be challenging when an agent is delayed, because this typically creates conflicts with other agents. So, we need to

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Prefer watching on @YouTube while scrolling through the comment section? We get it: http://youtu.be/b1Pvt072wKQ?si=NGPgm30ur1WFtKQS

DGX agent

Google AI shared a post on X (formerly Twitter) directing followers to watch their content on YouTube, suggesting their video is also available on that platform for viewers who prefer that experience.

model-releasesgoogle-ai--x
13 Apr 2026
Model Releases

Pretrain-then-Adapt: Uncertainty-Aware Test-Time Adaptation for Text-based Person Search

DGX agent

arXiv:2604.08598v1 Announce Type: cross Abstract: Text-based person search faces inherent limitations due to data scarcity, driven by stringent privacy constraints and the high cost of manual annotati

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Provable Post-Training Quantization: Theoretical Analysis of OPTQ and Qronos

DGX agent

arXiv:2508.04853v2 Announce Type: replace-cross Abstract: Post-training quantization (PTQ) has become a crucial tool for reducing the memory and compute costs of modern deep neural networks, including

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

QARIMA: A Quantum Approach To Classical Time Series Analysis

DGX agent

arXiv:2604.08277v2 Announce Type: replace-cross Abstract: We present a quantum-inspired ARIMA methodology that integrates quantum-assisted lag discovery with fixed-configuration variational quantum ci

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

QoS-QoE Translation with Large Language Model

DGX agent

arXiv:2604.08703v1 Announce Type: cross Abstract: QoS-QoE translation is a fundamental problem in multimedia systems because it characterizes how measurable system and network conditions affect user-p

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

QuanBench+: A Unified Multi-Framework Benchmark for LLM-Based Quantum Code Generation

DGX agent

arXiv:2604.08570v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for code generation, yet quantum code generation is still evaluated mostly within single frameworks

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Quantisation Reshapes the Metacognitive Geometry of Language Models

DGX agent

arXiv:2604.08976v1 Announce Type: new Abstract: We report that model quantisation restructures domain-level metacognitive efficiency in LLMs rather than degrading it uniformly. Evaluating Llama-3-8B-I

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

R2G: A Multi-View Circuit Graph Benchmark Suite from RTL to GDSII

DGX agent

arXiv:2604.08810v1 Announce Type: new Abstract: Graph neural networks (GNNs) are increasingly applied to physical design tasks such as congestion prediction and wirelength estimation, yet progress is

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

RADSeg: Unleashing Parameter and Compute Efficient Zero-Shot Open-Vocabulary Segmentation Using Agglomerative Models

DGX agent

arXiv:2511.19704v2 Announce Type: replace Abstract: Open-vocabulary semantic segmentation (OVSS) underpins many vision and robotics tasks that require generalizable semantic understanding. Existing ap

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

RansomTrack: A Hybrid Behavioral Analysis Framework for Ransomware Detection

DGX agent

arXiv:2604.08739v1 Announce Type: cross Abstract: Ransomware poses a serious and fast-acting threat to critical systems, often encrypting files within seconds of execution. Research indicates that ran

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Realism character reference is on the way. Please fill out the survey below to sign up and get notified when it launches. https://links.comf…

DGX agent

ComfyUI is developing a realism character reference feature and is collecting user interest through a survey prior to its official launch. Users can sign up via the provided survey link to receive not

model-releasescomfyui--x
13 Apr 2026
Model Releases

Reasoning in a Combinatorial and Constrained World: Benchmarking LLMs on Natural-Language Combinatorial Optimization

DGX agent

arXiv:2602.02188v2 Announce Type: replace Abstract: While large language models (LLMs) have shown strong performance in math and logic reasoning, their ability to handle combinatorial optimization (CO

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences

DGX agent

arXiv:2602.11354v2 Announce Type: replace Abstract: The literature has witnessed an emerging interest in AI agents for automated assessment of scientific papers. Existing benchmarks focus primarily on

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation

DGX agent

arXiv:2510.17640v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable performance on complex tasks through imitation learning in recent robotic man

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Retrieval Augmented Classification for Confidential Documents

DGX agent

arXiv:2604.08628v1 Announce Type: cross Abstract: Unauthorized disclosure of confidential documents demands robust, low-leakage classification. In real work environments, there is a lot of inflow and

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Revisiting Image Manipulation Localization under Realistic Manipulation Scenarios

DGX agent

arXiv:2509.20006v3 Announce Type: replace Abstract: With the large models easing the labor-intensive manipulation process, image manipulations in today's real scenarios often entail a complex manipula

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Robust Reasoning Benchmark

DGX agent

arXiv:2604.08571v1 Announce Type: cross Abstract: While Large Language Models (LLMs) achieve high performance on standard mathematical benchmarks, their underlying reasoning processes remain highly ov

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning

DGX agent

arXiv:2604.09452v1 Announce Type: cross Abstract: Safety guarantees are a prerequisite to the deployment of reinforcement learning (RL) agents in safety-critical tasks. Often, deployment environments

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

SAGE: A Service Agent Graph-guided Evaluation Benchmark

DGX agent

arXiv:2604.09285v1 Announce Type: new Abstract: The development of Large Language Models (LLMs) has catalyzed automation in customer service, yet benchmarking their performance remains challenging. Ex

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Scaling model size is hitting diminishing returns. The real gains are in orchestration. Our Co-Founder & Co-CEO @yshoham makes the case in a…

DGX agent

Scaling model size is hitting diminishing returns. The real gains are in orchestration. Our Co-Founder & Co-CEO @yshoham makes the case in a rare long-form profile by @Calcalistech today. The man tryi

model-releasesai21-labs--x
13 Apr 2026
Model Releases

SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment

DGX agent

arXiv:2604.08988v1 Announce Type: new Abstract: Current LLM-based agents demonstrate strong performance in episodic task execution but remain constrained by static toolsets and episodic amnesia, faili

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

See, Hear, and Understand: Benchmarking Audiovisual Human Speech Understanding in Multimodal Large Language Models

DGX agent

arXiv:2512.02231v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) are expected to jointly interpret vision, audio, and language, yet existing video benchmarks rarely a

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

See you this Wednesday at the Ollama Gemma Meetup! 💎

DGX agent

Ollama announced a meetup focused on Gemma, Google's open-weight language model family, scheduled for Wednesday. The event likely brought together developers and AI enthusiasts to explore running Gemm

model-releasesollama--x
13 Apr 2026
Model Releases

Seeing is Believing: Robust Vision-Guided Cross-Modal Prompt Learning under Label Noise

DGX agent

arXiv:2604.09532v1 Announce Type: cross Abstract: Prompt learning is a parameter-efficient approach for vision-language models, yet its robustness under label noise is less investigated. Visual conten

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Semantic Rate-Distortion for Bounded Multi-Agent Communication: Capacity-Derived Semantic Spaces and the Communication Cost of Alignment

DGX agent

arXiv:2604.09521v1 Announce Type: cross Abstract: When two agents of different computational capacities interact with the same environment, they need not compress a common semantic alphabet differentl

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

SenBen: Sensitive Scene Graphs for Explainable Content Moderation

DGX agent

arXiv:2604.08819v1 Announce Type: cross Abstract: Content moderation systems classify images as safe or unsafe but lack spatial grounding and interpretability: they cannot explain what sensitive behav

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Sentiment Classification of Gaza War Headlines: A Comparative Analysis of Large Language Models and Arabic Fine-Tuned BERT Models

DGX agent

arXiv:2604.08566v1 Announce Type: new Abstract: This study examines how different artificial intelligence architectures interpret sentiment in conflict-related media discourse, using the 2023 Gaza War

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

SessionIntentBench: A Multi-task Inter-session Intention-shift Modeling Benchmark for E-commerce Customer Behavior Understanding

DGX agent

arXiv:2507.20185v2 Announce Type: replace Abstract: Session history is a common way of recording user interacting behaviors throughout a browsing activity with multiple products. For example, if an us

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

SiMing-Bench: Evaluating Procedural Correctness from Continuous Interactions in Clinical Skill Videos

DGX agent

arXiv:2604.09037v1 Announce Type: cross Abstract: Current video benchmarks for multimodal large language models (MLLMs) focus on event recognition, temporal ordering, and long-context recall, but over

model-releasesarxiv-cs-cl
13 Apr 2026
← Previous
1…445446447448449…461
Next →