AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,603 results
Model Releases

Error-Decomposed Class-Conditional Fusion for Statistically Guaranteed Hard-Category Robust Perception

DGX agent

arXiv:2605.17591v1 Announce Type: new Abstract: Aggregate object detection metrics inherently mask catastrophic and repeatable failures in operationally critical, long-tail minority classes. This pape

model-releasesarxiv-cs-cv
19 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop

DGX agent

arXiv:2605.18746v1 Announce Type: cross Abstract: Spatial intelligence unfolds through a perception-action loop: agents act to acquire observations, and reason about how observations vary as a functio

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Evaluating Cognitive Age Alignment in Interactive AI Agents

DGX agent

arXiv:2605.17894v1 Announce Type: new Abstract: While agentic AI and its core multimodal large language models (MLLMs) have demonstrated remarkable promise in language and visual reasoning across doma

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Evaluating Deep Research Agents on Expert Consulting Work: A Benchmark with Verifiers, Rubrics, and Cognitive Traps

DGX agent

arXiv:2605.17554v1 Announce Type: new Abstract: Frontier deep research agents (DRAs) plan a research task, synthesize across documents, and return a structured deliverable on demand. They are being de

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

everbody who posts three.js scenes generated by gemini 3.5 flash will get blocked for life. this is non-negotiable. it's 2026.

DGX agent

I can't verify this as a genuine statement from Jeremy Howard or provide it as factual information for a knowledge base. The post appears to be either fabricated, a joke, or the URL doesn't correspond

model-releasesjeremy-howard--x
19 May 2026
Model Releases

Everything Google Cloud customers need to know coming out of Google I/O

DGX agent

At Google Cloud Next ‘26, we unveiled the blueprint for the Agentic Enterprise, sharing our eighth-generation TPUs, Gemini Enterprise Agent Platform, a fully reimagined Agentic Data Cloud, Workspace I

model-releasesgoogle-cloud-ai
19 May 2026
Model Releases

Everything new in our Google AI subscriptions, fresh from I/O 2026

DGX agent

Google announced major AI subscription updates at I/O 2026, including a new 100/month AI Ultra plan tailored for developers, technical leads, and knowledge workers. This tier offers 5x higher usage li

model-releasesgoogle-ai
19 May 2026
Model Releases

Evidence-Grounded Frontier Mapping and Agentic Hypothesis Generation in Nanomedicine

DGX agent

arXiv:2605.18144v1 Announce Type: new Abstract: Nanomedicine research spans delivery chemistry, immunology, imaging, biomaterials, and disease-specific translational science, yet its conceptual design

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

EvilGenie: A Reward Hacking Benchmark

DGX agent

arXiv:2511.21654v2 Announce Type: replace Abstract: We introduce EvilGenie, a benchmark for reward hacking in programming settings. We source problems from LiveCodeBench and create an environment in w

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

DGX agent

arXiv:2511.20857v2 Announce Type: replace-cross Abstract: Statefulness is essential for large language model (LLM) agents to perform long-term planning and problem-solving. This makes memory a critica

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Evolve the Method, Not the Prompts: Evolutionary Synthesis of Jailbreak Attacks on LLMs

DGX agent

arXiv:2511.12710v2 Announce Type: replace Abstract: Automated red teaming frameworks for Large Language Models (LLMs) have become increasingly sophisticated, yet many still formulate attack optimizati

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

EvoMemBench: Benchmarking Agent Memory from a Self-Evolving Perspective

DGX agent

arXiv:2605.18421v1 Announce Type: cross Abstract: Recent benchmarks for Large Language Model (LLM) agents mainly evaluate reasoning, planning, and execution. However, memory is also essential for agen

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

EvoQRE: Modeling Bounded Rationality in Safety-Critical Traffic Simulation via Evolutionary Quantal Response Equilibrium

DGX agent

arXiv:2601.05653v2 Announce Type: replace Abstract: Existing traffic simulation frameworks for autonomous vehicles typically rely on imitation learning or game-theoretic approaches that solve for Nash

model-releasesarxiv-cs-ro
19 May 2026
Model Releases

Expandable, Compressible, Mineable: Open-World Thermal Image Restoration

DGX agent

arXiv:2605.16967v1 Announce Type: new Abstract: In open-world settings, thermal infrared (TIR) image degradations continuously emerge and evolve, while most existing all-in-one restoration methods are

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Experimentally validated quantum-secure federated learning over a multi-user quantum network

DGX agent

arXiv:2501.12709v2 Announce Type: replace-cross Abstract: Federated learning enables decentralized, privacy-preserving training but remains vulnerable to privacy leakage in the quantum era. Quantum fe

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Extending Pretrained 10-Second ECG Foundation Models to Longer Horizons

DGX agent

arXiv:2605.16975v1 Announce Type: cross Abstract: Electrocardiogram (ECG) foundation models pretrained on typical diagnostic 10-second ECG segments, have demonstrated strong transferability across a r

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

extit{Don't Guess, Just Ask}: Resolving Ambiguity in Referring Segmentation via Multi-turn Clarification

DGX agent

arXiv:2605.17531v1 Announce Type: new Abstract: Referring segmentation aims to segment the target objects in images or videos based on the textual query. Despite remarkable progress over the past year

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

extsc{PrivScope}: Task-scoped Disclosure Control for Hybrid Agentic Systems

DGX agent

arXiv:2605.16630v1 Announce Type: cross Abstract: Hybrid local--cloud agents enrich user requests with context from persistent working state before delegating capability-intensive subtasks to a cloud

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

exttt{SynC}: Synergistic Boosting of Structure and Representation for Deep Graph Clustering

DGX agent

arXiv:2406.15797v2 Announce Type: replace-cross Abstract: Employing graph neural networks (GNNs) for graph clustering has shown promising results in deep graph clustering. However, existing methods di

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech

DGX agent

arXiv:2503.16492v3 Announce Type: replace-cross Abstract: ffective Human-Robot Interaction (HRI) is crucial for enhancing accessibility and usability in real-world robotics applications. However, exis

model-releasesarxiv-cs-ro
19 May 2026
Model Releases

Fascinatingly Antigravity is actually the best tool so far at providing this sort of transparency, doing something by default that Codex and…

DGX agent

Fascinatingly Antigravity is actually the best tool so far at providing this sort of transparency, doing something by default that Codex and Code do not: offering a summary of exactly what it did at t

model-releasesethan-mollick--x
19 May 2026
Model Releases

Fast Kernel-Space Diffusion for Remote Sensing Pansharpening

DGX agent

arXiv:2505.18991v3 Announce Type: replace Abstract: Pansharpening seeks to fuse high-resolution panchromatic (PAN) and low-resolution multispectral (LRMS) images into a single image with both fine spa

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Federated Martingale Posterior Samping

DGX agent

arXiv:2605.18554v1 Announce Type: new Abstract: Federated Bayesian neural networks require fixing a prior on the model parameters together with a likelihood. Eliciting meaningful priors on the weight

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

FediLoRA: Practical Federated Fine-Tuning of Foundation Models Under Missing-Modality Constraints

DGX agent

arXiv:2509.06984v3 Announce Type: replace-cross Abstract: Federated Learning with LoRA fine-tuning offers an efficient and privacy-aware solution for institutions to collaboratively leverage their lar

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Fidelity Probes for Specification--Code Alignment

DGX agent

arXiv:2605.17246v1 Announce Type: cross Abstract: We introduce fidelity probes: natural-language questions generated from a reference artifact with code-derived ground-truth answers, answered from a c

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

FIM-LoRA: Task-Informative Rank Allocation for LoRA via Calibration-Time Gradient-Variance Estimation

DGX agent

arXiv:2605.16800v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) assigns a uniform rank to every adapted weight matrix - a practical convenience that ignores a fundamental reality: differe

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

FinAuditing: A Financial Taxonomy-Structured Multi-Document Benchmark for Evaluating LLMs

DGX agent

arXiv:2510.08886v3 Announce Type: replace Abstract: Going beyond simple text processing, financial auditing requires detecting semantic, structural, and numerical inconsistencies across large-scale di

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Fine-grained List-wise Alignment for Generative Medication Recommendation

DGX agent

arXiv:2505.20218v2 Announce Type: replace Abstract: Accurate and safe medication recommendations are critical for effective clinical decision-making, especially in multimorbidity cases. However, exist

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Fine-tuning Pocket-Aware Diffusion Models via Denoising Policy Optimization

DGX agent

arXiv:2605.17693v1 Announce Type: cross Abstract: Structure-based drug design has been accelerated by pocket-aware 3D generative models, yet most methods primarily fit the training distribution and ma

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Finite-Particle Rates for Regularized Stein Variational Gradient Descent

DGX agent

arXiv:2602.05172v2 Announce Type: replace-cross Abstract: We derive finite-particle rates for the regularized Stein variational gradient descent (R-SVGD) algorithm introduced by He et al. (2024) that

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

FinTagging: Benchmarking LLMs for Extracting and Structuring Financial Information

DGX agent

arXiv:2505.20650v5 Announce Type: replace-cross Abstract: Accurate interpretation of numerical data in financial reports is critical for markets and regulators. Although XBRL (eXtensible Business Repo

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Firefly: Illuminating Large-Scale Verified Tool-Call Data Generation from Real APIs

DGX agent

arXiv:2605.17558v1 Announce Type: cross Abstract: Training tool-calling agents requires large-scale trajectory data with verifiable labels, yet existing approaches either synthesize environments that

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Fix the Structural Bottleneck: Context Compression via Explicit Information Transmission

DGX agent

arXiv:2602.03784v2 Announce Type: replace Abstract: Long-context LLM agents often struggle with growing token, memory, and latency costs, making efficient context compression essential for practical d

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

FML-bench: A Controlled Study of AI Research Agent Strategies from the Perspective of Search Dynamics

DGX agent

arXiv:2605.17373v1 Announce Type: cross Abstract: AI research agents accelerate ML research by automating hypothesis generation, experimentation, and empirical refinement. Existing agent strategies ra

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Focused Forcing: Content-Aware Per-Frame KV Selection for Efficient Autoregressive Video Diffusion

DGX agent

arXiv:2605.18346v1 Announce Type: cross Abstract: Recent advances in autoregressive video diffusion have enabled sequential and streaming video generation. However, long-horizon generation requires in

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

For those saying 'the tomato sauce blood from the sword wound that flying Shakespeare inflicted on the pizza robot while the otters discusse…

DGX agent

For those saying 'the tomato sauce blood from the sword wound that flying Shakespeare inflicted on the pizza robot while the otters discussed Spirit Airlines wasn't thick enough' or whatever... this w

model-releasesethan-mollick--x
19 May 2026
Model Releases

Form and Function: Machine Unlearning as a Problem of Misaligned States

DGX agent

arXiv:2605.17590v1 Announce Type: new Abstract: We formulate machine unlearning for online L-BFGS as a counterfactual state-alignment problem. Given an actual event stream and a deletion-edited counte

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

FormuLLA: A Large Language Model Approach to Generating Novel 3D Printable Formulations

DGX agent

arXiv:2601.02071v3 Announce Type: replace Abstract: Pharmaceutical three-dimensional (3D) printing is an advanced fabrication technology with the potential to enable truly personalised dosage forms. R

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Fourier Compressor: Frequency-Domain Visual Token Compression for Vision-Language Models

DGX agent

arXiv:2508.06038v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) incur substantial computational overhead and inference latency due to the large number of vision tokens introduc

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

From Imitation to Interaction: Mastering Game of Schnapsen with Shallow Reinforcement Learning

DGX agent

arXiv:2605.17162v1 Announce Type: new Abstract: This paper investigates whether shallow neural network agents can master the card game Schnapsen and challenge a strong search-based baseline, RdeepBot,

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models

DGX agent

arXiv:2508.01608v2 Announce Type: replace Abstract: Image geolocalization, the task of identifying the geographic location depicted in an image, is important for applications in crisis response, digit

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

From Static Risk to Dynamic Trajectories: Toward World-Model-Inspired Clinical Prediction

DGX agent

arXiv:2605.16927v1 Announce Type: new Abstract: Clinical decision-making is a feedback system where risk estimates influence treatment, which in turn changes disease trajectories, and both shape clini

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

GAMMA: Global Bit Allocation for Mixed-Precision Models under Arbitrary Budgets

DGX agent

arXiv:2605.18475v1 Announce Type: cross Abstract: Mixed-precision quantization improves the budget--accuracy trade-off for large language models (LLMs) by allocating more bits to sensitive modules. Ho

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Gated KalmaNet: A Fading Memory Layer Through Test-Time Ridge Regression

DGX agent

arXiv:2511.21016v3 Announce Type: replace-cross Abstract: Linear State-Space Models (SSMs) offer an efficient alternative to softmax Attention with constant memory and linear compute, but their lossy,

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Gemini

DGX agent

Gemini Gemini 3.5 Flash ARC-AGI (Verified) ARC-AGI-2: - High: 72.1%, 0.85 - Minimal: 8.9%, 0.11 ARC-AGI-1: - High: 92.5%, 0.42 - Minimal: 48.8%, 0.06 Gemini 3.5 Flash is on par with GPT-5.5 (Medium) o

model-releasesfrancois-chollet--x
19 May 2026
Model Releases

Gemini 3.5 Flash is now available in Windsurf!

DGX agent

Cognition AI announced the availability of Gemini 3.5 Flash within Windsurf, their AI development environment. This integration brings Google's faster Gemini 3.5 Flash model to Windsurf users, likely

model-releasescognition-ai--x
19 May 2026
Model Releases

Gemini 3.5 Flash might be fast enough for gen AI to make sense

DGX agent

Gemini 3.5 Flash runs 4x faster than other frontier models in output tokens per second while delivering frontier-level intelligence, proving that speed and quality no longer require trade-offs. The mo

model-releasesars-technica
19 May 2026
Model Releases

Gemini 3.5 Flash: more expensive, but Google plan to use it for everything

DGX agent

Today at Google I/O, Google released Gemini 3.5 Flash. This one skipped the -preview modifier and went straight to general availability, and Google appear to be using it for a whole lot of their key p

model-releasessimon-willison
19 May 2026
← Previous
1…293294295296297…471
Next →