AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,595 results
9 Jun 2026

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP

Model ReleasesDGX agent

arXiv:2505.11189v3 Announce Type: replace Abstract: Large language models (LLMs) can amplify misinformation, undermining societal goals such as the UN SDGs. We study three documented drivers of misinf

Can we stabilize an inverted pendulum with feedback from a time-of-flight camera?

Model ReleasesDGX agent

arXiv:2606.09237v1 Announce Type: new Abstract: Time-of-flight cameras are popular in robotics for providing direct depth information while being compact, inexpensive, and robust to lighting condition

Can You Trust What You See? Human and AI Detection of Synthetic Legal Evidence

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.07613v1 Announce Type: cross Abstract: Visual evidence has long been treated as a reliable form of legal proof, but advances in artificial intelligence (AI) are undermining that assumption.

CATPO: Critique-Augmented Tree Policy Optimization

Model ReleasesDGX agent

arXiv:2606.08346v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a dominant paradigm for improving the reasoning capabilities of large language models

Causal Agent Replay: Counterfactual Attribution for LLM-Agent Failures

Model ReleasesDGX agent

arXiv:2606.08275v1 Announce Type: cross Abstract: When an LLM agent fails -- issues a refund it should not have, calls the wrong tool, leaks data -- existing tooling answers what happened (observabili

Causal Longitudinal Prior-Fitted Networks for Counterfactual Outcome Prediction

Model ReleasesDGX agent

arXiv:2606.05797v2 Announce Type: replace Abstract: Longitudinal treatment decisions from multivariate time-series data require predicting potential outcomes under future treatment sequences in the pr

Chain of Flow: ECG-Conditioned 4D Cardiac Cine Generation from Patient-Specific Anatomical Anchor

Model ReleasesDGX agent

arXiv:2602.22919v2 Announce Type: replace Abstract: Cardiac cine magnetic resonance imaging (MRI) is central to functional cardiac assessment, yet a full current cine sequence may not always be direct

CHIMERA-Bench: A Benchmark Dataset for Epitope-Specific Antibody Design

Model ReleasesDGX agent

arXiv:2603.13431v3 Announce Type: replace-cross Abstract: Computational antibody design has seen rapid methodological progress, with dozens of deep generative methods proposed in the past three years,

ChinaHeritaQA: A Culturally-Grounded Visual Question Answering Dataset for World Heritage Sites in China

Model ReleasesDGX agent

arXiv:2606.08959v1 Announce Type: new Abstract: We introduce ChinaHeritaQA, a multimodal benchmark dataset for evaluating the cultural reasoning abilities of vision-language models (VLMs) on UNESCO Wo

CHROMA: Detecting AI-Generated Images through Inter-Channel Color-Space Correlations

Model ReleasesDGX agent

arXiv:2606.08864v1 Announce Type: new Abstract: The rapid adoption of diffusion and large-scale generative models has made it increasingly challenging to distinguish synthetic imagery from real photog

ChronoPhyBench: Do MLLMs Truly Understand the World or Merely Exploit Language Priors?

Model ReleasesDGX agent

arXiv:2606.07962v1 Announce Type: new Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated remarkable proficiency in open-world reasoning and understanding. Howe

CLASP: Language-Driven Robot Skill Selection and Composition using Task-Parameterized Learning

Model ReleasesDGX agent

arXiv:2606.08169v1 Announce Type: cross Abstract: Enabling robots to understand and execute tasks from natural language commands while maintaining data efficiency remains challenging. Foundation model

Claude Code-Driving Scenario Mining for the Argoverse 2 Challenge

Model ReleasesDGX agent

arXiv:2606.09180v1 Announce Type: new Abstract: We present our submission to the CVPR 2026 Argoverse 2 Scenario Mining Challenge. Our system uses a four-stage pipeline: (1) autonomous code generation

Claude Fable 5 and new AI safety fables

Model ReleasesDGX agent

This article discusses Claude Fable 5, likely exploring Anthropic's latest version of their AI model and examining new fables or narratives related to AI safety concepts. The piece probably analyzes h

Claude Fable 5: Available on Google Cloud

Model ReleasesDGX agent

Claude Fable 5, Anthropic’s latest frontier model, is now generally available on Google Cloud. This launch is the latest proof point of our ongoing commitment to bring the industry's latest models str

Claude Fable 5 hands-on: impressive results working on complex projects, like building an interactive isochrone map or a data analysis tool in just 9.5 hours (Ethan Mollick/One Useful Thing)

Model ReleasesDGX agent

Ethan Mollick / One Useful Thing: Claude Fable 5 hands-on: impressive results working on complex projects, like building an interactive isochrone map or a data analysis tool in just 9.5 hours — Claude

Claude Fable 5 is now available in Cursor. It sets a new state of the art on CursorBench at 72.9%, 8 points above the previous best.

Model ReleasesDGX agent

Claude Fable 5 is now available for use in the Cursor IDE, achieving a new state-of-the-art performance on CursorBench with a score of 72.9%, an 8-point improvement over the previous best result. This

Claude Fable 5 is now available in Devin. Fable 5 earns the #1 spot on FrontierCode, our benchmark for real-world engineering tasks that gra…

Model ReleasesDGX agent

Claude Fable 5 has been integrated into Devin and achieved the top ranking on FrontierCode, Cognition AI's benchmark for real-world engineering tasks. This integration represents an advancement in Dev

Claude Fable 5 is now available on Databricks, fully governed through Unity AI Gateway

Model ReleasesDGX agent

Claude Fable 5 is now available as a model option on the Databricks platform, integrated with Unity AI Gateway to provide governance controls for enterprise users. This integration allows organization

Claude Fable 5 is now supported for use in Hermes Agent via Nous Portal! The first 500 new users get one month free access to the Plus plan …

Model ReleasesDGX agent

Nous Research announced support for Claude Fable 5 integration with Hermes Agent through the Nous Portal, offering new users a promotional one-month free trial of the Plus plan for the first 500 signu

Claude Fable 5 now available on AI Gateway

Model ReleasesDGX agent

Claude Fable 5 is now available through Vercel's AI Gateway, expanding the model options developers can access through the platform. This announcement likely details how developers can integrate and u

Claude Mythos went from “too dangerous to release” to publicly available (with some extra guard rails) in two months. And y’all fell for Ant…

Model ReleasesDGX agent

Gary Marcus critiques Anthropic's rapid shift in positioning Claude from a model deemed too dangerous for public release to one made widely available with safety measures, suggesting this represents i

@claudeai please buy more data centers asap

Model ReleasesDGX agent

Jerry Liu's post urges Anthropic to rapidly expand its data center infrastructure to support Claude AI's growing demand and operational needs. The post reflects concerns about scaling computational re

Closing the Sim-to-Real Gap: An Evaluation Framework for Autonomous Cyber Defense Configuration of Commercial EDR

Model ReleasesDGX agent

arXiv:2606.08168v1 Announce Type: cross Abstract: Leading commercial endpoint detection and response (EDR) products have shifted from operator-configured rule sets to multi-component systems where aut

Coarse-to-Fine Hierarchical Alignment for UAV-based Human Detection using Diffusion Models

Model ReleasesDGX agent

arXiv:2512.13869v3 Announce Type: replace Abstract: Training object detectors demands extensive, task-specific annotations, yet this requirement becomes impractical in UAV-based human detection due to

CodeTaste: Can LLMs Generate Human-Level Code Refactorings?

Model ReleasesDGX agent

arXiv:2603.04177v2 Announce Type: replace-cross Abstract: LLM coding agents can generate working code, but their solutions often accumulate complexity, duplication, and architectural debt. Human devel

@cohere Nice! Great to see another open source model released. 🙌

Model ReleasesDGX agent

Cohere announced the release of another open source model, receiving positive reception from the community. The post was shared on X (formerly Twitter) and highlights Cohere's continued contribution t

ComplexConstraints and Beyond: Expert Rubrics for RLVR

Model ReleasesDGX agent

arXiv:2606.09118v1 Announce Type: new Abstract: As LLM capabilities advance rapidly, the evaluation methods used to assess them increasingly lag behind. Traditional benchmarks relied on programmatic v

Component Ablation for Efficient Hybrid Language Model Architectures: Performance, Resilience, and Compression Implications

Model ReleasesDGX agent

arXiv:2603.22473v2 Announce Type: replace-cross Abstract: Hybrid language models combine softmax attention with linear-time sequence mechanisms such as state-space or linear-attention layers, but the

Conan-embedding-v3: Fusing Modality-Specific Models for Omni-Modal Embedding

Model ReleasesDGX agent

arXiv:2606.09331v1 Announce Type: cross Abstract: Omni-modal retrieval promises a single embedding space for text, image, video, document, and audio inputs, but building such a unified retriever is di

Condition-Gated Reasoning for Context-Dependent Biomedical Question Answering

Model ReleasesDGX agent

arXiv:2602.17911v3 Announce Type: replace-cross Abstract: Current biomedical question answering (QA) systems often assume that medical knowledge applies uniformly, yet real-world clinical reasoning is

Conditional Normalizing Flows for Forward and Backward Joint State and Parameter Estimation

Model ReleasesDGX agent

arXiv:2601.07013v2 Announce Type: replace-cross Abstract: Traditional filtering algorithms for state estimation -- such as classical Kalman filtering, unscented Kalman filtering, and particle filters

Context Rot in AI-Assisted Software Development: Repurposing Documentation Consistency for AI Configuration Artifacts

Model ReleasesDGX agent

arXiv:2606.09090v1 Announce Type: cross Abstract: Developers increasingly provide AI coding assistants with persistent context through configuration files such as CLAUDE.md, AGENTS.md, and .cursorrule

ContextShift: A Controlled Benchmark for Context Dependence in Object Detection

Model ReleasesDGX agent

arXiv:2606.09495v1 Announce Type: new Abstract: Modern object detectors achieve strong performance on standard benchmarks, yet their robustness to contextual variation remains insufficiently understoo

Convergence Bound and Critical Batch Size of Muon Optimizer

Model ReleasesDGX agent

arXiv:2507.01598v5 Announce Type: replace Abstract: Muon, a recently proposed optimizer that leverages the inherent matrix structure of neural network parameters, has demonstrated strong empirical per

Convolutional Sparse Coding via the Locally Competitive Algorithm on Loihi 2

Model ReleasesDGX agent

arXiv:2606.08584v1 Announce Type: new Abstract: Sparse coding provides a principled framework for signal representation by expressing an input as a linear combination of only a small number of basis f

Correcting Mean Bias in Text Embeddings: A Refined Renormalization with Training-Free Improvements on MMTEB

Model ReleasesDGX agent

arXiv:2511.11041v2 Announce Type: replace-cross Abstract: We find that current sentence-embedding models produce outputs with a consistent bias: every embedding e decomposes as ilde e + mu, where the

Correlation Is Not Enough: Embedding Human Metadata for Individual Causal Discovery

Model ReleasesDGX agent

arXiv:2606.09672v1 Announce Type: new Abstract: Ask a pretrained biomedical language model whether 'cortisol 28 ug/dL' and 'stock-market volatility' are related, and it returns a cosine similarity of

Counterfactual Transport Flows for Offline Conservative Trajectory Refinement

Model ReleasesDGX agent

arXiv:2606.09115v1 Announce Type: new Abstract: Offline reinforcement learning (RL) offers a path to policy improvement from logged data alone, using historical returns or other measurable outcomes as

CoVEBench: Can Video Editing Models Handle Complex Instructions?

Model ReleasesDGX agent

arXiv:2606.08415v1 Announce Type: cross Abstract: While recent text-guided video editing models excel at elementary tasks (e.g., style transfer, object insertion), real-world user requests are highly

CRANE: Knowledge Editing for Reasoning MLLMs

Model ReleasesDGX agent

arXiv:2606.09033v1 Announce Type: new Abstract: The emergence of reasoning multimodal large language models (MLLMs), which generate explicit chain-of-thought (CoT) reasoning before producing answers,

Cross-View Urban Traffic Dataset: Drone-Supervised Ground Truth for Monocular Bird's-Eye View Localization

Model ReleasesDGX agent

arXiv:2606.07708v1 Announce Type: cross Abstract: We introduce a dataset and benchmark for cross-view urban traffic perception built from synchronized ego-centric bicycle videos and aerial drone video

CTS-Bench: Benchmarking Graph Coarsening Trade-offs for GNNs in Clock Tree Synthesis

Model ReleasesDGX agent

arXiv:2602.19330v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) are increasingly explored for physical design analysis in Electronic Design Automation, particularly for modeling Clock

Curvature-Guided LoRA: Matching Full Fine-Tuning in Function Space

Model ReleasesDGX agent

arXiv:2603.29824v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning methods such as LoRA enable efficient adaptation of large pretrained models, but often lag behind full fine-tuning i

DAL-PCQA: Enabling Distortion-Level and Language-Driven Reasoning for Point Cloud Quality Assessment

Model ReleasesDGX agent

arXiv:2606.07938v1 Announce Type: new Abstract: Point Cloud Quality Assessment (PCQA) methods typically predict scalar Mean Opinion Scores (MOS), which quantify overall perceptual degradation but do n

Data Synthesis and Parameter-Efficient Fine-Tuning for Low-Resource NMT: A Case Study on Q'eqchi' Mayan

Model ReleasesDGX agent

arXiv:2606.09767v1 Announce Type: cross Abstract: Neural machine translation for digitally low-resource Indigenous languages is often hindered by extreme data scarcity, prompting reliance on extractiv

Datadog launches more than 100 features at DASH to push autonomous AI ops

Model ReleasesDGX agent

Observability and security platform company Datadog Inc. today unveiled more than 100 new capabilities at its annual DASH 2026 conference, headlined by a major expansion of its Bits AI agents that the

De novo molecular generation with optical property preconditioning at the token level

Model ReleasesDGX agent

arXiv:2606.08221v1 Announce Type: new Abstract: Designing OLED molecules with targeted optical properties remains challenging due to the scarcity of high-quality data and the limited reliability of co

Dear @MeekMill Fable 5 just came out!

Model ReleasesDGX agent

Dear @MeekMill Fable 5 just came out! I feel like they can make my claude smarter who can help em do that ... or what is the smartest ai program available to the people? Because the things I am learni

Decision-Aware Memory Cards: Counterfactual-Inspired Context Selection and Compression for Tool-Using LLM Agents

Model ReleasesDGX agent

arXiv:2606.08151v1 Announce Type: new Abstract: Tool-using LLM agents often fail not because relevant text is absent, but because decisive evidence is not selected, compressed, or surfaced at action t

Declarative Outcome-Conformant Synthesis: Exact, Closed-Form Specification Satisfaction and a Conformance Benchmark

Model ReleasesDGX agent

arXiv:2606.08736v1 Announce Type: new Abstract: We study a capability the dominant paradigm in synthetic tabular data does not provide: exact satisfaction of a declared analytical outcome with no sour

Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models

Model ReleasesDGX agent

arXiv:2606.09142v1 Announce Type: cross Abstract: Egocentric vision offers a first-person view of human perception and decision making, yet its potential for traffic-safety prediction remains underexp

Decomposable Neuro Symbolic Regression

Model ReleasesDGX agent

arXiv:2511.04124v3 Announce Type: replace Abstract: Symbolic regression (SR) models complex systems by discovering mathematical expressions that capture underlying relationships in observed data. Howe

Deep Tree Tensor Networks

Model ReleasesDGX agent

arXiv:2502.09928v2 Announce Type: replace-cross Abstract: Originating in quantum physics, tensor networks (TNs) have been widely adopted as exponential machines and parametric decomposers for recognit

DeepMine-Mamba: Mitigating Information Dilution in Mamba-Based State Space Models for Document Image Binarization

Model ReleasesDGX agent

arXiv:2606.08781v1 Announce Type: new Abstract: Document image binarization aims to separate foreground text from degraded backgrounds while preserving thin, broken, and low-contrast strokes. Although

DeepSeek Inference Analysis Day 0 GB200 NVL72, Blackwell, AMD MI355X Huawei performance analysis Performance every day over time for the pas…

Model ReleasesDGX agent

DeepSeek Inference Analysis Day 0 GB200 NVL72, Blackwell, AMD MI355X Huawei performance analysis Performance every day over time for the past 44 days DeepSeekV4 1.6T Day 0 to Day 43 Performance Over T

Defending Against Malicious Finetuning by Scaling Train-time Adversarial Attacks

Model ReleasesDGX agent

arXiv:2606.07970v1 Announce Type: cross Abstract: Current open-weight large language models (LLMs) are prone to malicious finetuning attacks, which could compromise the safety alignment of LLMs with o

Dendrograms of Mixing Measures for Softmax-Gated Gaussian Mixture of Experts: Consistency Without Model Sweeps

Model ReleasesDGX agent

arXiv:2510.12744v2 Announce Type: replace-cross Abstract: We develop a unified statistical framework for softmax-gated Gaussian mixture of experts (SGMoE) that addresses three long-standing obstacles

Detecting and containing AI-powered threats with Google Security Operations agents

Model ReleasesDGX agent

To defend against the growing range of AI-accelerated threat actors, organizations need to be able to respond faster to outpace the adversary.Recently, we announced Google AI Threat Defense, an automa

Developing Distance-Aware Physics-Constrained Probabilistic Frameworks for Industrial Prognostics

Model ReleasesDGX agent

arXiv:2512.08499v3 Announce Type: replace-cross Abstract: Development of reliable and physically interpretable probabilistic frameworks for industrial prognostics remain nascent, and existing literatu

← Previous
1…154155156157158…377
Next →