AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,332 results
Model Releases

Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps

DGX agent

arXiv:2604.19565v1 Announce Type: cross Abstract: Hallucinations in Speech Large Language Models (SpeechLLMs) pose significant risks, yet existing detection methods typically rely on gold-standard out

model-releasesarxiv-cs-ai
22 Apr 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Detection of T-shirt Presentation Attacks in Face Recognition Systems

DGX agent

arXiv:2604.19365v1 Announce Type: new Abstract: Face recognition systems are often used for biometric authentication. Nevertheless, it is known that without any protective measures, face recognition s

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

Disparities In Negation Understanding Across Languages In Vision-Language Models

DGX agent

arXiv:2604.18942v1 Announce Type: new Abstract: Vision-language models (VLMs) exhibit affirmation bias: a systematic tendency to select positive captions ('X is present') even when the correct descrip

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Do Agents Dream of Root Shells? Partial-Credit Evaluation of LLM Agents in Capture The Flag Challenges

DGX agent

arXiv:2604.19354v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly proposed for autonomous cybersecurity tasks, but their capabilities in realistic offensive settings r

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Do LLMs Game Formalization? Evaluating Faithfulness in Logical Reasoning

DGX agent

arXiv:2604.19459v1 Announce Type: new Abstract: Formal verification guarantees proof validity but not formalization faithfulness. For natural-language logical reasoning, where models construct axiom s

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Does Self-Consistency Improve the Recall of Encyclopedic Knowledge?

DGX agent

arXiv:2604.19395v1 Announce Type: new Abstract: While self-consistency is known to improve performance on symbolic reasoning, its effect on the recall of encyclopedic knowledge is unclear due to a lac

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

DP-FlogTinyLLM: Differentially private federated log anomaly detection using Tiny LLMs

DGX agent

arXiv:2604.19118v1 Announce Type: cross Abstract: Modern distributed systems generate massive volumes of log data that are critical for detecting anomalies and cyber threats. However, in real world se

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

DUALVISION: RGB-Infrared Multimodal Large Language Models for Robust Visual Reasoning

DGX agent

arXiv:2604.18829v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved impressive performance on visual perception and reasoning tasks with RGB imagery, yet they remain

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

DW-Bench: Benchmarking LLMs on Data Warehouse Graph Topology Reasoning

DGX agent

arXiv:2604.18964v1 Announce Type: new Abstract: This paper introduces DW-Bench, a new benchmark that evaluates large language models (LLMs) on graph-topology reasoning over data warehouse schemas, exp

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

EfficientPENet: Real-Time Depth Completion from Sparse LiDAR via Lightweight Multi-Modal Fusion

DGX agent

arXiv:2604.18790v1 Announce Type: new Abstract: Depth completion from sparse LiDAR measurements and corresponding RGB images is a prerequisite for accurate 3D perception in robotic systems. Existing m

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

Emotion-Cause Pair Extraction in Conversations via Semantic Decoupling and Graph Alignment

DGX agent

arXiv:2604.19547v1 Announce Type: new Abstract: Emotion-Cause Pair Extraction in Conversations (ECPEC) aims to identify the set of causal relations between emotion utterances and their triggering caus

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Enabling the agentic enterprise: business and industry agents arrive in Gemini Enterprise

DGX agent

We are officially moving past the era of one-size-fits-all AI. Enterprises today require highly specialized, role-specific tools to drive real productivity — but they cannot afford to sacrifice securi

model-releasesgoogle-cloud-ai
22 Apr 2026
Model Releases

Energy-Weighted Flow Matching: Unlocking Continuous Normalizing Flows for Efficient and Scalable Boltzmann Sampling

DGX agent

arXiv:2509.03726v2 Announce Type: replace-cross Abstract: Sampling from unnormalized target distributions, e.g. Boltzmann distributions mu_{ext{target}}(x) propto exp(-E(x)/T), is fundamental to many

model-releasesarxiv-cs-lg
22 Apr 2026
Model Releases

Enjoyed the read? If you have deep experience in ML frameworks (training or inference) and love working on problems like these, our team is …

DGX agent

Enjoyed the read? If you have deep experience in ML frameworks (training or inference) and love working on problems like these, our team is hiring! ML Systems Engineer, Frameworks & Tooling: https://j

model-releasescohere--x
22 Apr 2026
Model Releases

Environmental Sound Deepfake Detection Using Deep-Learning Framework

DGX agent

arXiv:2604.19652v1 Announce Type: cross Abstract: In this paper, we propose a deep-learning framework for environmental sound deepfake detection (ESDD) -- the task of identifying whether the sound sce

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Evaluating Answer Leakage Robustness of LLM Tutors against Adversarial Student Attacks

DGX agent

arXiv:2604.18660v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in education, yet their default helpfulness often conflicts with pedagogical principles. Prior work

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Exabeam extends Agent Behavior Analytics to Google Cloud’s agent ecosystem

DGX agent

Security intelligence and management solutions company Exabeam Inc. today announced new Exabeam Agent Behavior Analytics capabilities for agents built with Google Cloud’s Agent Development Kit and an

model-releasessiliconangle
22 Apr 2026
Model Releases

Excellent research by Conway Zhu, Ali Edalati, and Zewen Shen. Read the blog here: https://cohere.com/blog/vllm-integration-and-quality-reco…

DGX agent

Cohere researchers Conway Zhu, Ali Edalati, and Zewen Shen published research on vLLM integration and quality recommendations. The blog post, shared via Cohere's official X account, likely discusses o

model-releasescohere--x
22 Apr 2026
Model Releases

Fast and Robust Diffusion Posterior Sampling for MR Image Reconstruction Using the Preconditioned Unadjusted Langevin Algorithm

DGX agent

arXiv:2512.05791v2 Announce Type: replace-cross Abstract: Purpose: The Unadjusted Langevin Algorithm (ULA) in combination with diffusion models can generate high quality MRI reconstructions with uncer

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

FedProxy: Federated Fine-Tuning of LLMs via Proxy SLMs and Heterogeneity-Aware Fusion

DGX agent

arXiv:2604.19015v1 Announce Type: cross Abstract: Federated fine-tuning of Large Language Models (LLMs) is obstructed by a trilemma of challenges: protecting LLMs intellectual property (IP), ensuring

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Fine-tuning DeepSeek-OCR-2 for Molecular Structure Recognition

DGX agent

arXiv:2604.03476v2 Announce Type: replace-cross Abstract: Optical Chemical Structure Recognition (OCSR) is critical for converting 2D molecular diagrams from printed literature into machine-readable f

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Fine-Tuning Small Reasoning Models for Quantum Field Theory

DGX agent

arXiv:2604.18936v1 Announce Type: cross Abstract: Despite the growing application of Large Language Models (LLMs) to theoretical physics, there is little academic exploration into how domain-specific

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

FluentAvatar: Flicker-Free Talking-Head Animation via Phoneme-Guided Autoregressive Modeling

DGX agent

arXiv:2509.12052v3 Announce Type: replace Abstract: Current talking-head generation has gradually shifted from GAN-based methods to diffusion-based paradigms, achieving remarkable progress in visual f

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

For the 'small test' they've modified their docs to remove mention of Claude Code in Claude Pro: https://support.claude.com/en/articles/1114…

DGX agent

For the 'small test' they've modified their docs to remove mention of Claude Code in Claude Pro: https://support.claude.com/en/articles/11145838-using-claude-code-with-your-max-plan It's been a shock

model-releasesjeremy-howard--x
22 Apr 2026
Model Releases

Four-Axis Decision Alignment for Long-Horizon Enterprise AI Agents

DGX agent

arXiv:2604.19457v1 Announce Type: new Abstract: Long-horizon enterprise agents make high-stakes decisions (loan underwriting, claims adjudication, clinical review, prior authorization) under lossy mem

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

From Experience to Skill: Multi-Agent Generative Engine Optimization via Reusable Strategy Learning

DGX agent

arXiv:2604.19516v1 Announce Type: new Abstract: Generative engines (GEs) are reshaping information access by replacing ranked links with citation-grounded answers, yet current Generative Engine Optimi

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

From Natural Language to Executable Narsese: A Neuro-Symbolic Benchmark and Pipeline for Reasoning with NARS

DGX agent

arXiv:2604.18873v1 Announce Type: new Abstract: Large language models (LLMs) are highly capable at language generation, but they remain unreliable when reasoning requires explicit symbolic structure,

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

From Proof to Program: Characterizing Tool-Induced Reasoning Hallucinations in Large Language Models

DGX agent

arXiv:2511.10899v2 Announce Type: replace Abstract: Tool-augmented Language Models (TaLMs) can invoke external tools to solve problems beyond their parametric capacity. However, it remains unclear whe

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Gated Memory Policy

DGX agent

arXiv:2604.18933v1 Announce Type: cross Abstract: Robotic manipulation tasks exhibit varying memory requirements, ranging from Markovian tasks that require no memory to non-Markovian tasks that depend

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Gemini Cloud Assist: Proactive cloud operations that work for you, even before you ask

DGX agent

Today at Google Cloud Next, we are unveiling a more proactive Gemini Cloud Assist, our AI-assisted cloud operations platform. This update shifts your Google Cloud operations from manual workflows to a

model-releasesgoogle-cloud-ai
22 Apr 2026
Model Releases

Gemma 4 VLA Demo on Jetson Orin Nano Super

DGX agent

This article demonstrates running Gemma 4, Google's open-weight language model, on NVIDIA's Jetson Orin Nano Super edge computing device. It likely covers the model's capabilities, performance metrics

model-releaseshugging-face
22 Apr 2026
Model Releases

GenerativeMPC: VLM-RAG-guided Whole-Body MPC with Virtual Impedance for Bimanual Mobile Manipulation

DGX agent

arXiv:2604.19522v1 Announce Type: new Abstract: Bimanual mobile manipulation requires a seamless integration between high-level semantic reasoning and safe, compliant physical interaction - a challeng

model-releasesarxiv-cs-ro
22 Apr 2026
Model Releases

GeoLaux: A Benchmark for Evaluating MLLMs' Geometry Performance on Long-Step Problems Requiring Auxiliary Lines

DGX agent

arXiv:2508.06226v2 Announce Type: replace Abstract: Geometry problem solving (GPS) poses significant challenges for Multimodal Large Language Models (MLLMs) in diagram comprehension, knowledge applica

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Get more from speculative decoding in MoE models https://cohere.link/Et2rbsB

DGX agent

Speculative decoding is a technique that accelerates language model inference by using a smaller draft model to generate candidate tokens, which are then verified by a larger model, reducing latency w

model-releasescohere--x
22 Apr 2026
Model Releases

Google launches AI research agents powered by Gemini 3.1 Pro

DGX agent

Google LLC has released two artificial intelligence agents that can generate research reports about user-specified topics. The agents made their debut on Tuesday. Deep Research and Deep Research Max a

model-releasessiliconangle
22 Apr 2026
Model Releases

Google Meet will take AI notes for in-person meetings too

DGX agent

Google's AI meeting notetaker is no longer limited to Google Meets - Gemini can also generate summaries and transcripts of in-person meetings now, as well as meetings on Zoom and Microsoft Teams, as f

model-releasesthe-verge-ai
22 Apr 2026
Model Releases

Google puts Gemini Enterprise at the heart of the new agentic taskforce for enterprise automation

DGX agent

Google Cloud is on a mission to accelerate the adoption of artificial intelligence agents across enterprise computing environments, paving the way for a new era where AI can automate many of the most

model-releasessiliconangle
22 Apr 2026
Model Releases

Google says 75% of new code created inside the company is now generated by AI and reviewed by human engineers, up from 50% last fall (Hugh Langley/Business Insider)

DGX agent

Hugh Langley / Business Insider: Google says 75% of new code created inside the company is now generated by AI and reviewed by human engineers, up from 50% last fall — - Three-quarters of new code at

model-releasestechmeme
22 Apr 2026
Model Releases

Google says Meet's Gemini-powered Take Notes for Me feature can now be used for in-person meetings and adds support for Teams and Zoom (Abner Li/9to5Google)

DGX agent

Abner Li / 9to5Google: Google says Meet's Gemini-powered Take Notes for Me feature can now be used for in-person meetings and adds support for Teams and Zoom — Besides Workspace Intelligence, Google a

model-releasestechmeme
22 Apr 2026
Model Releases

GRAFT: Geometric Refinement and Fitting Transformer for Human Scene Reconstruction

DGX agent

arXiv:2604.19624v1 Announce Type: new Abstract: Reconstructing physically plausible 3D human-scene interactions (HSI) from a single image currently presents a trade-off: optimization based methods off

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

GRASPrune: Global Gating for Budgeted Structured Pruning of Large Language Models

DGX agent

arXiv:2604.19398v1 Announce Type: new Abstract: Large language models (LLMs) are expensive to serve because model parameters, attention computation, and KV caches impose substantial memory and latency

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Guys, I am absolutely astounded. The Qwen 3.6 27b is like a jump to Qwen 4 from Qwen 27B 3.5. I just did a full suite of front end design te…

DGX agent

Guys, I am absolutely astounded. The Qwen 3.6 27b is like a jump to Qwen 4 from Qwen 27B 3.5. I just did a full suite of front end design tests and agentic benchmarks, made entirely by it. VERDICT: Th

model-releasesclem-delangue--x
22 Apr 2026
Model Releases

HalluAudio: A Comprehensive Benchmark for Hallucination Detection in Large Audio-Language Models

DGX agent

arXiv:2604.19300v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have recently achieved strong performance across various audio-centric tasks. However, hallucination, where models

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

HarDBench: A Benchmark for Draft-Based Co-Authoring Jailbreak Attacks for Safe Human-LLM Collaborative Writing

DGX agent

arXiv:2604.19274v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as co-authors in collaborative writing, where users begin with rough drafts and rely on LLMs to compl

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Harmful Intent as a Geometrically Recoverable Feature of LLM Residual Streams

DGX agent

arXiv:2604.18901v1 Announce Type: cross Abstract: Harmful intent is geometrically recoverable from large language model residual streams: as a linear direction in most layers, and as angular deviation

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

HarmoniDiff-RS: Training-Free Diffusion Harmonization for Satellite Image Composition

DGX agent

arXiv:2604.19392v1 Announce Type: new Abstract: Satellite image composition plays a critical role in remote sensing applications such as data augmentation, disaste simulation, and urban planning. We p

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

Has anyone seen ANY official communication from Anthropic or an Anthropic staff member about the fact that the checkbox for Claude Code on P…

DGX agent

Has anyone seen ANY official communication from Anthropic or an Anthropic staff member about the fact that the checkbox for Claude Code on Pro is back to being checked again, or is the only evidence t

model-releasessimon-willison--x
22 Apr 2026
Model Releases

Has Automated Essay Scoring Reached Sufficient Accuracy? Deriving Achievable QWK Ceilings from Classical Test Theory

DGX agent

arXiv:2604.19131v1 Announce Type: new Abstract: Automated essay scoring (AES) is commonly evaluated on public benchmarks using quadratic weighted kappa (QWK). However, because benchmark labels are ass

model-releasesarxiv-cs-ai
22 Apr 2026
← Previous
1…400401402403404…466
Next →