AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
Model Releases

As believers of open research, we are disappointed to see Anthropic silently degrading Fable 5 for AI development 'Any topic related to buil…

DGX agent

As believers of open research, we are disappointed to see Anthropic silently degrading Fable 5 for AI development 'Any topic related to building pretraining pipelines, distributed training infrastruct

model-releasesyann-lecun--x
10 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

ASyMOB: Algebraic Symbolic Mathematical Operations Benchmark

DGX agent

arXiv:2505.23851v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied to symbolic mathematics, yet existing evaluations often conflate pattern memorization wi

model-releasesarxiv-cs-ai
10 Jun 2026
Research

AuRA: Internalizing Audio Understanding into LLMs as LoRA

DGX agent

arXiv:2606.11033v1 Announce Type: cross Abstract: Recent efforts to extend large language models (LLMs) to speech inputs typically rely on cascaded ASR-LLM pipelines, end-to-end speech-language models

researcharxiv-cs-ai
10 Jun 2026
Model Releases

Beyond APIs: Probing the Limits of MLLMs in Physical Tool Use

DGX agent

arXiv:2606.10803v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) excel at utilizing digital APIs and increasingly serve as the 'brain' of embodied AI, instructing robots to i

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

CoCoSI: Collaborative Cognitive Map Construction for Spatial Intelligence

DGX agent

arXiv:2606.10401v1 Announce Type: new Abstract: Spatial intelligence is a key frontier for multimodal large language models (MLLMs), enabling them to reason about the physical world from visual experi

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

ComBench: A Benchmark for Rigorous Proof Reasoning and Constructive Realization in Olympiad-Level Combinatorics

DGX agent

arXiv:2606.10479v1 Announce Type: new Abstract: Combinatorics is central to Olympiad-level mathematical problem solving, requiring deep discrete reasoning, creative constructions, and rigorous structu

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data

DGX agent

arXiv:2601.18026v2 Announce Type: replace Abstract: Language identification (LID) is a fundamental step in curating multilingual corpora. However, LID models still perform poorly for many languages, e

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Compile Once, Differentiate Everywhere: A Differentiable Meta-Circular Interpreter

DGX agent

arXiv:2606.09930v1 Announce Type: cross Abstract: The boundary between program execution and gradient-based optimization has long limited the use of code itself as a learnable scientific model. We pre

model-releasesarxiv-cs-lg
10 Jun 2026
Research

Cost-Aware Routing for Efficient Text-To-Image Generation

DGX agent

arXiv:2506.14753v3 Announce Type: replace Abstract: Diffusion models are well known for their ability to generate a high-fidelity image for an input prompt through an iterative denoising process. Unfo

researcharxiv-cs-cv
10 Jun 2026
Model Releases

Expert-Level Crisis Detection in Mental Health Conversations

DGX agent

arXiv:2606.10380v1 Announce Type: cross Abstract: Real-world crisis intervention is inherently conversational, yet existing research largely focuses on static texts.Real-world crisis intervention is i

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Fact-Augmented Lookahead Planning for LLM Agents

DGX agent

arXiv:2506.09171v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly capable, but LLM agents still struggle to plan effectively in interactive, partially observable,

model-releasesarxiv-cs-ai
10 Jun 2026
Research

Human-AI Teaming Through the Lens of Calibration

DGX agent

arXiv:2606.10906v1 Announce Type: cross Abstract: We study models for human-AI teaming through the lens of statistical calibration. We assume the team consists of an AI model and human -- both of whic

researcharxiv-cs-ai
10 Jun 2026
Model Releases

Kwai Keye-VL-2.0 Technical Report

DGX agent

arXiv:2606.10651v1 Announce Type: new Abstract: We introduce Kwai Keye-VL-2.0-30B-A3B, an open-source Mixture-of-Experts (MoE) multimodal foundation model designed to advance long-video understanding

model-releasesarxiv-cs-cv
10 Jun 2026
Research

Linguistically Augmented Audio Speech Data (LinguAS)

DGX agent

arXiv:2606.10246v1 Announce Type: cross Abstract: Maliciously-created fake speech, including deepfaked and spoofed audio, is proliferating at an alarming rate, and detection models are racing to stay

researcharxiv-cs-ai
10 Jun 2026
Model Releases

LLM-as-a-Discriminator: When Synthetic Tables Still Look Real

DGX agent

arXiv:2606.09865v1 Announce Type: new Abstract: Privacy and data sharing are often in tension. Many organizations use synthetic data to reduce privacy risk and still share useful data. For tabular dat

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Local Is Not a Sufficient Privacy Boundary: Governing OS-Integrated On-Device AI

DGX agent

arXiv:2606.10173v1 Announce Type: cross Abstract: As AI systems move into operating systems, privacy no longer turns only on whether a model runs locally. A local assistant may assemble email, calenda

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

MIRAGE: A Polarity-Flipping Encoding Subspace in LLM Agents

DGX agent

arXiv:2606.10304v1 Announce Type: new Abstract: When LLM agents are coerced into covertly encoding sensitive data (Base64, ROT13, acrostic, synonym chains, and beyond), the resulting outputs evade out

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Optimal Post-Training Quantization Scales and Where to Find Them

DGX agent

arXiv:2606.10890v1 Announce Type: cross Abstract: Post-training quantization (PTQ) compresses large language models by mapping weights to low-bit representations. The scaling factor that defines the q

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Piper: A Programmable Distributed Training System

DGX agent

arXiv:2606.11169v1 Announce Type: cross Abstract: Large-scale model training increasingly relies on composing multiple parallelism strategies, such as data, pipeline, and expert parallelism, together

model-releasesarxiv-cs-ai
10 Jun 2026
Research

Pre-AF 13: An Interpretable Atrial Fibrillation Risk Score Mined from Discharge Reports

DGX agent

arXiv:2606.10725v1 Announce Type: cross Abstract: Background. Atrial fibrillation (AF) is the most prevalent cardiac arrhythmia and a major determinant of prognosis. Established AF risk scores rely on

researcharxiv-cs-cl
10 Jun 2026
Model Releases

Quoting Jeremy Howard

DGX agent

Easy solution to slow down recursive AI self improvement: The lab with the top-ranked model must agree THEY must not use it for working on frontier AI But everyone else should have access to it. By de

model-releasessimon-willison
10 Jun 2026
Research

RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty

DGX agent

arXiv:2602.12424v2 Announce Type: replace-cross Abstract: Benchmarks establish a standardized evaluation framework to systematically assess the performance of large language models (LLMs), facilitatin

researcharxiv-cs-ai
10 Jun 2026
Model Releases

SPACE: Source-free Proxy Anchor Concept Erasure for MLLMs

DGX agent

arXiv:2606.09868v1 Announce Type: cross Abstract: As Multimodal Large Language Models (MLLMs) face growing privacy risks and regulatory constraints, machine unlearning (MU) has emerged as a crucial so

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

The Interlocutor Effect: Why LLMs Leak More Personal Data to Agents Than Humans

DGX agent

arXiv:2606.09844v1 Announce Type: cross Abstract: Large Language Models (LLMs) alter their privacy behavior based on the perceived identity of their interlocutor. While safety mechanisms typically pre

model-releasesarxiv-cs-ai
10 Jun 2026
Research

VFUSE: Virulent Feature Understanding with Sparse autoEncoders

DGX agent

arXiv:2606.10080v1 Announce Type: cross Abstract: Generative models have shown remarkable progress in a variety of domains such as protein design, but such power enables the opaque generation of hazar

researcharxiv-cs-ai
10 Jun 2026
Model Releases

A Survey of Heterogeneous Graph Neural Networks for Cybersecurity Anomaly Detection

DGX agent

arXiv:2510.26307v3 Announce Type: replace-cross Abstract: Anomaly detection is a critical task in cybersecurity, where identifying insider threats, access violations, and coordinated attacks is essent

model-releasesarxiv-cs-lg
9 Jun 2026
Local Ai

Active Flow Expansion for Out-of-Distribution Discovery: from Theory to Molecules

DGX agent

arXiv:2606.08802v1 Announce Type: new Abstract: Standard flow and diffusion pre-training matches the distribution of available data (e.g., molecules), which often covers only a small fraction of the v

local-aiarxiv-cs-lg
9 Jun 2026
Model Releases

AutoMegaKernel: A Statically-Checked Agent Harness for Self-Retargeting Megakernel Synthesis

DGX agent

arXiv:2606.09682v1 Announce Type: new Abstract: AutoMegaKernel (AMK) compiles a HuggingFace Llama-family model into a single persistent cooperative CUDA kernel that runs the whole forward pass in one

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Causal Longitudinal Prior-Fitted Networks for Counterfactual Outcome Prediction

DGX agent

arXiv:2606.05797v2 Announce Type: replace Abstract: Longitudinal treatment decisions from multivariate time-series data require predicting potential outcomes under future treatment sequences in the pr

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

ChinaHeritaQA: A Culturally-Grounded Visual Question Answering Dataset for World Heritage Sites in China

DGX agent

arXiv:2606.08959v1 Announce Type: new Abstract: We introduce ChinaHeritaQA, a multimodal benchmark dataset for evaluating the cultural reasoning abilities of vision-language models (VLMs) on UNESCO Wo

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

End-to-End Training for Discrete Token LLM based TTS System

DGX agent

arXiv:2606.09234v1 Announce Type: cross Abstract: Recent state-of-the-art (SOTA) text-to-speech (TTS) systems typically adopt a cascaded pipeline consisting of a speech tokenizer, an autoregressive la

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ERBench: A Benchmark and Testsuite for Equation Discovery Algorithms

DGX agent

arXiv:2606.09276v1 Announce Type: new Abstract: Equation discovery aims to automate the discovery of scientific models in the form of mathematical equations from data. Technically, equation discovery

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting

DGX agent

arXiv:2606.09809v1 Announce Type: new Abstract: AI evaluation results are produced at scale but reported inconsistently across leaderboards, model cards, benchmark papers, and company blogs. The cost

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

GD-MIL: Grade-Disentangled Multiple Instance Learning for Multimodal Biochemical Recurrence Prediction in Prostate Cancer

DGX agent

arXiv:2606.09453v1 Announce Type: new Abstract: Biochemical recurrence (BCR) after radical prostatectomy is a critical endpoint in prostate cancer, yet risk stratification relies almost entirely on va

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Generalization in Nonlinear Least Squares via Learned Feature Geometry

DGX agent

arXiv:2606.08799v1 Announce Type: cross Abstract: We study the generalization of ridge-regularized nonlinear least-squares models via on-average algorithmic stability, deriving error bounds for local

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

GIScholarBench: Benchmarking LLM Overconfidence in GIS Research

DGX agent

arXiv:2606.08036v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in academic research workflows, but scholarly tasks require high factual precision and therefore ex

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

GRPO Does Not Close the Multi-Agent Coordination Gap

DGX agent

arXiv:2606.07845v1 Announce Type: cross Abstract: We measure how well current large language models coordinate as multiple agents sharing a common resource, using the dining philosophers problem as a

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

Guided Discovery of New Behaviors using Diffusion Policies

DGX agent

arXiv:2606.08743v1 Announce Type: new Abstract: Diffusion models have become a powerful tool for generative modeling in robotics, with diffusion policies excelling at modeling multimodal action-trajec

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

Hybrid Robustness Verification for Spatio-Temporal Neural Networks

DGX agent

arXiv:2606.09746v1 Announce Type: cross Abstract: With AI increasingly deployed in safety-critical systems, providing formal robustness guarantees for the underlying models is essential. Existing veri

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Language-based Trial and Error Falls Behind in the Era of Experience

DGX agent

arXiv:2601.21754v3 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in language-based agentic tasks, their applicability to unseen, nonlinguistic environments (e.g., symbolic

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Learning from Human Driving: A Human-in-the-Loop Online Behavior Cloning Framework for Autonomous Driving

DGX agent

arXiv:2606.08170v1 Announce Type: new Abstract: With the evolution of large foundation models (LFMs), data-driven autonomous driving has made significant strides. However, existing paradigms still fac

model-releasesarxiv-cs-ro
9 Jun 2026
Local Ai

Mean Teacher based SSL Framework for Indoor Localization Using Wi-Fi RSSI Fingerprinting

DGX agent

arXiv:2407.13303v2 Announce Type: replace Abstract: Conventional large-scale indoor localization based on Wi-Fi RSSI fingerprinting faces issues of time-consuming and labor-intensive labeled data coll

local-aiarxiv-cs-lg
9 Jun 2026
Applications

MedicalRec: Medical recommender system for image classification without retraining

DGX agent

arXiv:2606.07553v1 Announce Type: cross Abstract: The emergence of machine learning and deep learning has revolutionized the efficiency of diagnostic, therapeutic, and administrative systems in health

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

MMR-GRPO: Accelerating GRPO-Style Training through Diversity-Aware Reward Reweighting

DGX agent

arXiv:2601.09085v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) has become a standard approach for training mathematical reasoning models; however, its reliance on

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Now You (Still) See Me: Detecting Evasive Steganographic Payloads in LLMs

DGX agent

arXiv:2606.09411v1 Announce Type: cross Abstract: Large language models can be fine-tuned to encode prompt-borne secrets into fluent, seemingly benign outputs. This creates a steganographic exfiltrati

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

OmniCap-IF: Benchmarking and Improving Instruction Following Abilities for Omni-Video Captioning

DGX agent

arXiv:2606.08572v1 Announce Type: new Abstract: While Omni-modal Large Language Models (OLLMs) have demonstrated impressive capabilities in jointly processing audio and visual streams, their ability t

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Report: GKE Inference Gateway delivers up to 92% faster AI responses

DGX agent

As generative AI moves from experimental pilots to massive production environments, the efficiency of your infrastructure becomes the ultimate differentiator. One way to get the most out of it and min

model-releasesgoogle-cloud-ai
9 Jun 2026
Research

Rewrite to Translate, Translate to Reward: Reinforcement Learning for Source Rewriting in Machine Translation

DGX agent

arXiv:2606.08011v1 Announce Type: cross Abstract: Although directly prompting off-the-shelf Large Language Models (LLMs) to generate meaning-preserving source rewrites can effectively enhance Machine

researcharxiv-cs-ai
9 Jun 2026
← Previous
1…392393394395396…1326
Next →