AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
All
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,499 results
Model Releases

MCAP: Deployment-Time Layer Profiling for Memory-Constrained LLM Inference

DGX agent

arXiv:2604.21026v1 Announce Type: new Abstract: Deploying large language models to heterogeneous hardware is often constrained by memory, not compute. We introduce MCAP (Monte Carlo Activation Profili

model-releasesarxiv-cs-lg
24 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Measuring Opinion Bias and Sycophancy via LLM-based Coercion

DGX agent

arXiv:2604.21564v1 Announce Type: new Abstract: Large language models increasingly shape the information people consume: they are embedded in search, consulted for professional advice, deployed as age

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

mGRADE: Minimal Recurrent Gating Meets Delay Convolutions for Lightweight Sequence Modeling

DGX agent

arXiv:2507.01829v2 Announce Type: replace-cross Abstract: Multi-timescale sequence modeling relies on capturing both local fast dynamics and global slow context; yet, maintaining these capabilities un

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

MISTY: High-Throughput Motion Planning via Mixer-based Single-step Drifting

DGX agent

arXiv:2604.21489v1 Announce Type: cross Abstract: Multi-modal trajectory generation is essential for safe autonomous driving, yet existing diffusion-based planners suffer from high inference latency d

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Model page for more information and integrations: https://ollama.com/library/deepseek-v4-flash

DGX agent

Ollama announced DeepSeek-v4-flash, a lightweight variant of the DeepSeek-v4 model, now available in their model library for local deployment and integration. The model page provides documentation, us

model-releasesollama--x
24 Apr 2026
Model Releases

More of my notes on DeepSeek V4 - the really big news is the pricing: both DeepSeek-V4-Flash and DeepSeek-V4-Pro are the cheapest models in …

DGX agent

More of my notes on DeepSeek V4 - the really big news is the pricing: both DeepSeek-V4-Flash and DeepSeek-V4-Pro are the cheapest models in their categories while benchmarking close to the frontier mo

model-releasessimon-willison--x
24 Apr 2026
Model Releases

Multilingual and Domain-Agnostic Tip-of-the-Tongue Query Generation for Simulated Evaluation

DGX agent

arXiv:2604.21096v1 Announce Type: cross Abstract: Tip-of-the-Tongue (ToT) retrieval benchmarks have largely focused on English, limiting their applicability to multilingual information access. In this

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Musical Score Understanding Benchmark: Evaluating Large Language Models' Comprehension of Complete Musical Scores

DGX agent

arXiv:2511.20697v4 Announce Type: replace-cross Abstract: Understanding complete musical scores entails integrated reasoning over pitch, rhythm, harmony, and large-scale structure, yet the ability of

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

My first two TiKZ Sparks unicorns from DeepSeek v4. (Expert mode, from the DeepSeek site, which is supposed to be v4 Pro according to the re…

DGX agent

Ethan Mollick documents his first attempts at using DeepSeek v4's expert mode to generate TiKZ code for creating unicorn graphics, sharing results from the DeepSeek website's v4 Pro interface. The pos

model-releasesethan-mollick--x
24 Apr 2026
Model Releases

Navigating the Clutter: Waypoint-Based Bi-Level Planning for Multi-Robot Systems

DGX agent

arXiv:2604.21138v1 Announce Type: cross Abstract: Multi-robot control in cluttered environments is a challenging problem that involves complex physical constraints, including robot-robot collisions, r

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Necessity is the mother of kv-cache optimisation

DGX agent

Necessity is the mother of kv-cache optimisation I’m still amazed that DeepSeek, Kimi, and Qwen can train very strong LLMs with far fewer and often nerfed NVIDIA GPUs, or even Huawei chips. DeepSeek V

model-releasesemad-mostaque--x
24 Apr 2026
Model Releases

Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models

DGX agent

arXiv:2604.21896v1 Announce Type: new Abstract: This paper introduces a new paradigm for AI game programming, leveraging large language models (LLMs) to extend and operationalize Claude Shannon's taxo

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Neural surrogates for crystal growth dynamics with variable supersaturation: explicit vs. implicit conditioning

DGX agent

arXiv:2604.21753v1 Announce Type: cross Abstract: Simulations of crystal growth are performed by using Convolutional Recurrent Neural Network surrogate models, trained on a dataset of time sequences c

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

🐳Ollama is working to have DeepSeek-V4-Pro and Flash available on Ollama's cloud.

DGX agent

🐳Ollama is working to have DeepSeek-V4-Pro and Flash available on Ollama's cloud. 🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 Dee

model-releasesollama--x
24 Apr 2026
Model Releases

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

DGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

OmniFit: Multi-modal 3D Body Fitting via Scale-agnostic Dense Landmark Prediction

DGX agent

arXiv:2604.21575v1 Announce Type: new Abstract: Fitting an underlying body model to 3D clothed human assets has been extensively studied, yet most approaches focus on either single-modal inputs such a

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

On the Role of Preprocessing and Memristor Dynamics in Reservoir Computing for Image Classification

DGX agent

arXiv:2604.21602v1 Announce Type: cross Abstract: Reservoir computing (RC) is an emerging recurrent neural network architecture that has attracted growing attention for its low training cost and modes

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Open-H-Embodiment: A Large-Scale Dataset for Enabling Foundation Models in Medical Robotics

DGX agent

arXiv:2604.21017v1 Announce Type: cross Abstract: Autonomous medical robots hold promise to improve patient outcomes, reduce provider workload, democratize access to care, and enable superhuman precis

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

OpenAI GPT-5.5 now available on Databricks, fully-governed through Unity AI Gateway

DGX agent

OpenAI's GPT-5.5 model is now available on Databricks' platform with governance capabilities provided through Unity AI Gateway, enabling enterprises to deploy and manage the model within their data in

model-releasesdatabricks
24 Apr 2026
Model Releases

OpenEstimate: Evaluating LLMs on Reasoning Under Uncertainty with Real-World Data

DGX agent

arXiv:2510.15096v2 Announce Type: replace Abstract: Real-world settings where language models (LMs) are deployed -- in domains spanning healthcare, finance, and other forms of knowledge work -- requir

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

OptiVerse: A Comprehensive Benchmark towards Optimization Problem Solving

DGX agent

arXiv:2604.21510v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning, complex optimization tasks remain challenging, requiring domain knowledge and robus

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Nex…

DGX agent

Our teams have been busyyy! Here are some key updates from the past week: — @GoogleCloud unveiled a suite of AI innovations at our Cloud Next event, including our eighth generation TPUs (TPUt for infe

model-releasesgoogle-ai--x
24 Apr 2026
Model Releases

pi gives you wings wherever you are

DGX agent

pi gives you wings wherever you are This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚‍♀️ Qwen3.6 27B running inside of Pi coding agent via Llama.cpp on the MacBook Pro Fo

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

pi + @ollama + gemma4 + @p0 - each very useful individually, but so much better when composed together!

DGX agent

pi + @ollama + gemma4 + @p0 - each very useful individually, but so much better when composed together! We built a completely free CLI agent with @badlogicgames's Pi agent, @ollama (Gemma 4), and Para

model-releasesollama--x
24 Apr 2026
Model Releases

Pretrain Where? Investigating How Pretraining Data Diversity Impacts Geospatial Foundation Model Performance

DGX agent

arXiv:2604.21104v1 Announce Type: new Abstract: New geospatial foundation models introduce a new model architecture and pretraining dataset, often sampled using different notions of data diversity. Pe

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

PREVENT-JACK: Context Steering for Swarms of Long Heavy Articulated Vehicles

DGX agent

arXiv:2604.21337v1 Announce Type: new Abstract: In this paper, we aim to extend the traditional point-mass-like robot representation in swarm robotics and instead study a swarm of long Heavy Articulat

model-releasesarxiv-cs-ro
24 Apr 2026
Model Releases

Principled Evaluation with Human Labels: One Rater at a Time and Rater Equivalence

DGX agent

arXiv:2106.01254v3 Announce Type: replace Abstract: In many classification tasks, there is no definitive ground truth, only human judgments that may disagree. We address two challenges that arise in s

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Process Supervision via Verbal Critique Improves Reasoning in Large Language Models

DGX agent

arXiv:2604.21611v1 Announce Type: cross Abstract: Inference-time scaling for LLM reasoning has focused on three axes: chain depth, sample breadth, and learned step-scorers (PRMs). We introduce a fourt

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

【Qwen-3.6-27B × llama.cpp】生成速度10倍の革命的スピード! Qwen-3.6-27Bで生成速度が約10倍の136.75 t/sに到達する驚異の手法が話題です!🚀 llama.cppの「ngram-mod」という投機的デコード(Speculative D…

DGX agent

【Qwen-3.6-27B × llama.cpp】生成速度10倍の革命的スピード! Qwen-3.6-27Bで生成速度が約10倍の136.75 t/sに到達する驚異の手法が話題です!🚀 llama.cppの「ngram-mod」という投機的デコード(Speculative Decoding)を活用。過去の出力パターンを利用して次に来る言葉を予測し、追加のビデオメモリをほぼ消費せずに高速化を実現し

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

RailVQA: A Benchmark and Framework for Efficient Interpretable Visual Cognition in Automatic Train Operation

DGX agent

arXiv:2603.27112v2 Announce Type: replace Abstract: As Automatic Train Operation (ATO) advances toward GoA4 and beyond, it increasingly depends on efficient, reliable cab-view visual perception and de

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was g…

DGX agent

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was going to be more difficult and that I would be 'fighting' wit

model-releasesdair-ai--x
24 Apr 2026
Model Releases

RealRoute: Dynamic Query Routing System via Retrieve-then-Verify Paradigm

DGX agent

arXiv:2604.20860v1 Announce Type: cross Abstract: Despite the success of Retrieval-Augmented Generation (RAG) in grounding LLMs with external knowledge, its application over heterogeneous sources (e.g

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Reasoning About Traversability: Language-Guided Off-Road 3D Trajectory Planning

DGX agent

arXiv:2604.21249v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) enable high-level semantic reasoning for end-to-end autonomous driving, particularly in unstructured environments, e

model-releasesarxiv-cs-ro
24 Apr 2026
Model Releases

Rectified Schrodinger Bridge Matching for Few-Step Visual Navigation

DGX agent

arXiv:2604.05673v2 Announce Type: replace-cross Abstract: Visual navigation is a core challenge in Embodied AI, requiring autonomous agents to translate high-dimensional sensory observations into cont

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

ReFACT: A Benchmark for Scientific Confabulation Detection with Positional Error Annotations

DGX agent

arXiv:2509.25868v3 Announce Type: replace Abstract: The mechanisms underlying scientific confabulation in Large Language Models (LLMs) remain poorly understood. We introduce ReFACT (Reddit False And C

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Reinforcing 3D Understanding in Point-VLMs via Geometric Reward Credit Assignment

DGX agent

arXiv:2604.21160v1 Announce Type: new Abstract: Point-Vision-Language Models promise to empower embodied agents with executable spatial reasoning, yet they frequently succumb to geometric hallucinatio

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Reinforcing privacy reasoning in LLMs via normative simulacra from fiction

DGX agent

arXiv:2604.20904v1 Announce Type: cross Abstract: Information handling practices of LLM agents are broadly misaligned with the contextual privacy expectations of their users. Contextual Integrity (CI)

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Remember o3 was only a year and a week ago! Also, only GPT-5.5 seemed to take the 'evolution' piece seriously and change the setting rather …

DGX agent

I cannot provide an accurate summary for this entry as the text appears incomplete and lacks sufficient context. The post fragment references o3 (likely an AI model), GPT-5.5, and discusses timeline/e

model-releasesethan-mollick--x
24 Apr 2026
Model Releases

Retrofit: Continual Learning with Controlled Forgetting for Binary Security Detection and Analysis

DGX agent

arXiv:2511.11439v2 Announce Type: replace-cross Abstract: Binary security has increasingly relied on deep learning to reason about malware behavior and program semantics. However, the performance ofte

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Revealing Geography-Driven Signals in Zone-Level Claim Frequency Models: An Empirical Study using Environmental and Visual Predictors

DGX agent

arXiv:2604.21893v1 Announce Type: cross Abstract: Geographic context is often consider relevant to motor insurance risk, yet public actuarial datasets provide limited location identifiers, constrainin

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

RewardBench 2: Advancing Reward Model Evaluation

DGX agent

arXiv:2506.01937v2 Announce Type: replace Abstract: Reward models are used throughout the post-training of language models to capture nuanced signals from preference data and provide a training target

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Robust Test-time Video-Text Retrieval: Benchmarking and Adapting for Query Shifts

DGX agent

arXiv:2604.20851v1 Announce Type: cross Abstract: Modern video-text retrieval (VTR) models excel on in-distribution benchmarks but are highly vulnerable to real-world query shifts, where the distribut

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

SatSAM2: Motion-Constrained Video Object Tracking in Satellite Imagery using Promptable SAM2 and Kalman Priors

DGX agent

arXiv:2511.18264v3 Announce Type: replace Abstract: Existing satellite video tracking methods often struggle with generalization, requiring scenario-specific training to achieve satisfactory performan

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Scaling of Gaussian Kolmogorov--Arnold Networks

DGX agent

arXiv:2604.21174v1 Announce Type: cross Abstract: The Gaussian scale parameter (epsilon) is central to the behavior of Gaussian Kolmogorov--Arnold Networks (KANs), yet its role in deep edge-based arch

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

SCASeg: Strip Cross-Attention for Efficient Semantic Segmentation

DGX agent

arXiv:2411.17061v2 Announce Type: replace Abstract: The Vision Transformer (ViT) has achieved notable success in computer vision, with its variants widely validated across various downstream tasks, in

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

SCM: Sleep-Consolidated Memory with Algorithmic Forgetting for Large Language Models

DGX agent

arXiv:2604.20943v1 Announce Type: new Abstract: We present SCM (Sleep-Consolidated Memory), a research preview of a memory architecture for large language models that draws on neuroscientific principl

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Secure LLM Fine-Tuning via Safety-Aware Probing

DGX agent

arXiv:2505.16737v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable success across many applications, but their ability to generate harmful content raises s

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models

DGX agent

arXiv:2604.21523v1 Announce Type: cross Abstract: Large Vision-Language Models (VLMs) are increasingly used to evaluate outputs of other models, for image-to-text (I2T) tasks such as visual question a

model-releasesarxiv-cs-cl
24 Apr 2026
← Previous
1…394395396397398…469
Next →