AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,499 results
Model Releases

Separable Expert Architecture: Toward Privacy-Preserving LLM Personalization via Composable Adapters and Deletable User Proxies

DGX agent

arXiv:2604.21571v1 Announce Type: new Abstract: Current model training approaches incorporate user information directly into shared weights, making individual data removal computationally infeasible w

model-releasesarxiv-cs-ai
24 Apr 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Serialisation Strategy Matters: How FHIR Data Format Affects LLM Medication Reconciliation

DGX agent

arXiv:2604.21076v1 Announce Type: cross Abstract: Medication reconciliation at clinical handoffs is a high-stakes, error-prone process. Large language models are increasingly proposed to assist with t

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

SlideAgent: Hierarchical Agentic Framework for Multi-Page Visual Document Understanding

DGX agent

arXiv:2510.26615v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) extends large language models (LLMs) with external knowledge, but it must balance limited effective context, re

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

SocraticKG: Knowledge Graph Construction via QA-Driven Fact Extraction

DGX agent

arXiv:2601.10003v2 Announce Type: replace Abstract: Constructing Knowledge Graphs (KGs) from unstructured text provides a structured framework for knowledge representation and reasoning, yet current L

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

SparseGF: A Height-Aware Sparse Segmentation Framework with Context Compression for Robust Ground Filtering Across Urban to Natural Scenes

DGX agent

arXiv:2604.21356v1 Announce Type: new Abstract: High-quality digital terrain models derived from airborne laser scanning (ALS) data are essential for a wide range of geospatial analyses, and their gen

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Spatial Metaphors for LLM Memory: A Critical Analysis of the MemPalace Architecture

DGX agent

arXiv:2604.21284v1 Announce Type: new Abstract: MemPalace is an open-source AI memory system that applies the ancient method of loci (memory palace) spatial metaphor to organize long-term memory for l

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Spectral Embeddings Leak Graph Topology: Theory, Benchmark, and Adaptive Reconstruction

DGX agent

arXiv:2604.21094v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) excel on relational data, but standard benchmarks unrealistically assume the graph is centrally available. In practice, set

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Spud 🥔 and DeepSeek 🐳 V4 on the same day?! Is it Christmas? Here goes our night to bring the whale up 🔨

DGX agent

Fireworks AI announced the release or deployment of DeepSeek V4, a large language model, alongside another project or update called 'Spud' on the same day, with the team planning to work through the n

model-releasesfireworks-ai--x
24 Apr 2026
Model Releases

SQLyzr: A Comprehensive Benchmark and Evaluation Platform for Text-to-SQL

DGX agent

arXiv:2604.21214v1 Announce Type: cross Abstract: Text-to-SQL models have significantly improved with the adoption of Large Language Models (LLMs), leading to their increasing use in real-world applic

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Stealthy Backdoor Attacks against LLMs Based on Natural Style Triggers

DGX agent

arXiv:2604.21700v1 Announce Type: cross Abstract: The growing application of large language models (LLMs) in safety-critical domains has raised urgent concerns about their security. Many recent studie

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Strategic Heterogeneous Multi-Agent Architecture for Cost-Effective Code Vulnerability Detection

DGX agent

arXiv:2604.21282v1 Announce Type: cross Abstract: Automated code vulnerability detection is critical for software security, yet existing approaches face a fundamental trade-off between detection accur

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Structured Visual Narratives Undermine Safety Alignment in Multimodal Large Language Models

DGX agent

arXiv:2603.21697v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) extend text-only LLMs with visual reasoning, but also introduce new safety failure modes under visual

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

StyleID: A Perception-Aware Dataset and Metric for Stylization-Agnostic Facial Identity Recognition

DGX agent

arXiv:2604.21689v1 Announce Type: cross Abstract: Creative face stylization aims to render portraits in diverse visual idioms such as cartoons, sketches, and paintings while retaining recognizable ide

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Subject-level Inference for Realistic Text Anonymization Evaluation

DGX agent

arXiv:2604.21211v1 Announce Type: new Abstract: Current text anonymization evaluation relies on span-based metrics that fail to capture what an adversary could actually infer, and assumes a single dat

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

SurgViVQA: Temporally-Grounded Video Question Answering for Surgical Scene Understanding

DGX agent

arXiv:2511.03325v3 Announce Type: replace Abstract: Video Question Answering (VideoQA) in the surgical domain aims to enhance intraoperative understanding by enabling AI models to reason over temporal

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Symbolic Grounding Reveals Representational Bottlenecks in Abstract Visual Reasoning

DGX agent

arXiv:2604.21346v1 Announce Type: new Abstract: Vision--language models (VLMs) often fail on abstract visual reasoning benchmarks such as Bongard problems, raising the question of whether the main bot

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

SyMTRS: Benchmark Multi-Task Synthetic Dataset for Depth, Domain Adaptation and Super-Resolution in Aerial Imagery

DGX agent

arXiv:2604.21801v1 Announce Type: cross Abstract: Recent advances in deep learning for remote sensing rely heavily on large annotated datasets, yet acquiring high-quality ground truth for geometric, r

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Synthetic Data in Education: Empirical Insights from Traditional Resampling and Deep Generative Models

DGX agent

arXiv:2604.21031v1 Announce Type: cross Abstract: Synthetic data generation offers promise for addressing data scarcity and privacy concerns in educational technology, yet practitioners lack empirical

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

TEMA: Anchor the Image, Follow the Text for Multi-Modification Composed Image Retrieval

DGX agent

arXiv:2604.21806v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) is an important image retrieval paradigm that enables users to retrieve a target image using a multimodal query that cons

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Temporal Taskification in Streaming Continual Learning: A Source of Evaluation Instability

DGX agent

arXiv:2604.21930v1 Announce Type: new Abstract: Streaming Continual Learning (CL) typically converts a continuous stream into a sequence of discrete tasks through temporal partitioning. We argue that

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Tesla FSD is the first AI saving lives at scale....on real roads every single day Road accidents kill 1.19 million people every year - the #…

DGX agent

Tesla FSD is the first AI saving lives at scale....on real roads every single day Road accidents kill 1.19 million people every year - the #1 cause of death for ages 5–29 94%+ of crashes are caused by

model-releaseselon-musk--x
24 Apr 2026
Model Releases

The Coding Assistant Breakdown: More Tokens Please

DGX agent

This analysis examines the token consumption and economics of coding assistants, likely comparing different AI models' efficiency and cost-effectiveness for code generation tasks. The piece probably d

model-releasessemianalysis
24 Apr 2026
Model Releases

The Download: supercharged scams and studying AI healthcare

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. We’re in a new era of AI-driven scams When ChatGPT was release

model-releasesmit-tech-review
24 Apr 2026
Model Releases

The Feedback Hamiltonian is the Score Function: A Diffusion-Model Framework for Quantum Trajectory Reversal

DGX agent

arXiv:2604.21210v1 Announce Type: cross Abstract: In continuously monitored quantum systems, the feedback protocol of Garcia-Pintos, Liu, and Gorshkov reshapes the arrow of time: a Hamiltonian H_{meas

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

The First Challenge on Remote Sensing Infrared Image Super-Resolution at NTIRE 2026: Benchmark Results and Method Overview

DGX agent

arXiv:2604.21312v1 Announce Type: cross Abstract: This paper presents the NTIRE 2026 Remote Sensing Infrared Image Super-Resolution (x4) Challenge, one of the associated challenges of NTIRE 2026. The

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

The Newton–Schulz iteration coefficients optimized by DeepSeek-V4 are surprisingly strong: they effectively normalize all singular values to…

DGX agent

The Newton–Schulz iteration coefficients optimized by DeepSeek-V4 are surprisingly strong: they effectively normalize all singular values to 1. This matches our previous intuition: a well-balanced spe

model-releasesemad-mostaque--x
24 Apr 2026
Model Releases

The Path Not Taken: Duality in Reasoning about Program Execution

DGX agent

arXiv:2604.20917v1 Announce Type: cross Abstract: Large language models (LLMs) have shown remarkable capabilities across diverse coding tasks. However, their adoption requires a true understanding of

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

The Recurrent Transformer: Greater Effective Depth and Efficient Decoding

DGX agent

arXiv:2604.21215v1 Announce Type: new Abstract: Transformers process tokens in parallel but are temporally shallow: at position t, each layer attends to key-value pairs computed based on the previous

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

The Root Theorem of Context Engineering

DGX agent

arXiv:2604.20874v1 Announce Type: cross Abstract: Every system that maintains a large language model conversation beyond a single session faces two inescapable constraints: the context window is finit

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

The US CFTC sues New York, accusing the state of invading its authority to regulate prediction markets by filing lawsuits against Coinbase and Gemini (Jonathan Stempel/Reuters)

DGX agent

Jonathan Stempel / Reuters: The US CFTC sues New York, accusing the state of invading its authority to regulate prediction markets by filing lawsuits against Coinbase and Gemini — The U.S. Commodity F

model-releasestechmeme
24 Apr 2026
Model Releases

The week just gets better the mad men from China do it again China is hot Deepseek V4 Pro Let’s see how it is https://huggingface.co/deepsee…

DGX agent

DeepSeek V4 Pro is a new AI model release from Chinese AI company DeepSeek, announced via Hugging Face. The post expresses enthusiasm about the model's capabilities and performance, suggesting it repr

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

These pelicans are kind of angry looking! Left is deepseek-v4-flash, right is deepseek-v4-pro - both generated using OpenRouter via my LLM t…

DGX agent

These pelicans are kind of angry looking! Left is deepseek-v4-flash, right is deepseek-v4-pro - both generated using OpenRouter via my LLM tool 🚀 DeepSeek-V4 Preview is officially live & open-sourced!

model-releasessimon-willison--x
24 Apr 2026
Model Releases

Thinking Like a Botanist: Challenging Multimodal Language Models with Intent-Driven Chain-of-Inquiry

DGX agent

arXiv:2604.20983v1 Announce Type: cross Abstract: Vision evaluations are typically done through multi-step processes. In most contemporary fields, experts analyze images using structured, evidence-bas

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

This has been the most anticipated event in OSS and it doesn't disappoint. Congratulations to @deepseek_ai team. DSV4 Pro is now available o…

DGX agent

This has been the most anticipated event in OSS and it doesn't disappoint. Congratulations to @deepseek_ai team. DSV4 Pro is now available on @togethercompute and we'll be adding a lot of capacity beh

model-releasestogether-ai--x
24 Apr 2026
Model Releases

this is probably the most important piece of software of the decade next to vllm and sglang. i'm not joking.

DGX agent

this is probably the most important piece of software of the decade next to vllm and sglang. i'm not joking. llama.cpp at 100k stars now that 90% of the code worldwide is being written by AI agents, I

model-releasesjeremy-howard--x
24 Apr 2026
Model Releases

This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚‍♀️ Qwen3.6 27B running inside of Pi coding agent via Llama.…

DGX agent

This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚‍♀️ Qwen3.6 27B running inside of Pi coding agent via Llama.cpp on the MacBook Pro For non-trivial tasks on the @huggingf

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

Three reasons why DeepSeek’s new model matters

DGX agent

On Friday, Chinese AI firm DeepSeek released a preview of V4, its long-awaited new flagship model. Notably, the model can process much longer prompts than its last generation, thanks to a new design t

model-releasesmit-tech-review
24 Apr 2026
Model Releases

TimePre: Bridging Accuracy, Efficiency, and Stability in Probabilistic Time-Series Forecasting

DGX agent

arXiv:2511.18539v2 Announce Type: replace-cross Abstract: We propose TimePre, a simple framework that unifies the efficiency of Multilayer Perceptron (MLP)-based models with the distributional flexibi

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

To See the Unseen: on the Generalization Ability of Transformers in Symbolic Reasoning

DGX agent

arXiv:2604.21632v1 Announce Type: new Abstract: We investigate the ability of decoder-only transformer models to perform abstract symbolic reasoning; specifically solving propositional logic reasoning

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Tool Attention Is All You Need: Dynamic Tool Gating and Lazy Schema Loading for Eliminating the MCP/Tools Tax in Scalable Agentic Workflows

DGX agent

arXiv:2604.21816v1 Announce Type: new Abstract: The Model Context Protocol (MCP) has become a common interface for connecting large language model (LLM) agents to external tools, but its reliance on s

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Toward Efficient Membership Inference Attacks against Federated Large Language Models: A Projection Residual Approach

DGX agent

arXiv:2604.21197v1 Announce Type: new Abstract: Federated Large Language Models (FedLLMs) enable multiple parties to collaboratively fine-tune LLMs without sharing raw data, addressing challenges of l

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Towards Multimodal Active Learning: Efficient Learning with Limited Paired Data

DGX agent

arXiv:2510.03247v2 Announce Type: replace-cross Abstract: Active learning (AL) is a principled strategy to reduce annotation cost in data-hungry deep learning. However, existing AL algorithms focus al

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Towards Universal Tabular Embeddings: A Benchmark Across Data Tasks

DGX agent

arXiv:2604.21696v1 Announce Type: new Abstract: Tabular foundation models aim to learn universal representations of tabular data that transfer across tasks and domains, enabling applications such as t

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Transient Turn Injection: Exposing Stateless Multi-Turn Vulnerabilities in Large Language Models

DGX agent

arXiv:2604.21860v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into sensitive workflows, raising the stakes for adversarial robustness and safety. This pape

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

TRAVELFRAUDBENCH: A Configurable Evaluation Framework for GNN Fraud Ring Detection in Travel Networks

DGX agent

arXiv:2604.21093v1 Announce Type: cross Abstract: We introduce TravelFraudBench (TFG), a configurable benchmark for evaluating graph neural networks (GNNs) on fraud ring detection in travel platform g

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Trustworthy Clinical Decision Support Using Meta-Predicates and Domain-Specific Languages

DGX agent

arXiv:2604.21263v1 Announce Type: new Abstract: extbf{Background:} Regulatory frameworks for AI in healthcare, including the EU AI Act and FDA guidance on AI/ML-based medical devices, require clinical

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Try out Devin with the GPT-5.5 Agent Preview today: https://devin.ai

DGX agent

Cognition AI announced an early preview opportunity for Devin integrated with GPT-5.5 Agent capabilities, inviting users to test the combination at devin.ai. This likely represents an update to Devin'

model-releasescognition-ai--x
24 Apr 2026
Model Releases

Understanding and Mitigating Spurious Signal Amplification in Test-Time Reinforcement Learning for Math Reasoning

DGX agent

arXiv:2604.21327v1 Announce Type: cross Abstract: Test-time reinforcement learning (TTRL) always adapts models at inference time via pseudo-labeling, leaving it vulnerable to spurious optimization sig

model-releasesarxiv-cs-ai
24 Apr 2026
← Previous
1…395396397398399…469
Next →