AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlog
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,676 results
Safety

Speculative Decoding at Temperature Zero: A Scoped Safety-Invariance Screen with a 48,072-Sample Expansion

DGX agent

arXiv:2606.25097v1 Announce Type: new Abstract: Speculative decoding accelerates inference by letting a draft model propose tokens for a target model to verify, raising a concrete safety question: at

safetyarxiv-cs-lg
25 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Speech Codec Probing from Semantic and Phonetic Perspectives

DGX agent

arXiv:2603.10371v2 Announce Type: replace-cross Abstract: Speech tokenizers are essential for connecting speech to large language models (LLMs) in multimodal systems. Speech tokenizers are expected to

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Streaming-dLLM: Accelerating Diffusion LLMs via Suffix Pruning and Dynamic Decoding

DGX agent

arXiv:2601.17917v3 Announce Type: replace-cross Abstract: Diffusion Large Language Models (dLLMs) offer a compelling paradigm for natural language generation, leveraging parallel decoding and bidirect

researcharxiv-cs-cl
25 Jun 2026
Model Releases

V-Zero: Answer-Label-Free On-Policy Distillation with Contrastive Evidence Gating for Fine-Grained Visual Reasoning

DGX agent

arXiv:2606.25319v1 Announce Type: new Abstract: Fine-grained visual reasoning requires multimodal large language models (MLLMs) to identify task-relevant visual evidence and ground their reasoning in

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

VPA-Guard: Defending and Benchmarking Image-to-Video Generation Against Visual Prompt Attacks

DGX agent

arXiv:2606.25592v1 Announce Type: new Abstract: Recent advancements in Image-to-Video (I2V) generation have transformed input images from simple appearance references into interactive control interfac

model-releasesarxiv-cs-cv
25 Jun 2026
Safety

What Does It Mean to Break a Distillation Defense?

DGX agent

arXiv:2606.25059v1 Announce Type: cross Abstract: Black-box LLMs (accessible only via API) are vulnerable to distillation attacks, in which an attacker queries the model and trains a student on its ou

safetyarxiv-cs-ai
25 Jun 2026
Model Releases

What Does the Brain See? Multiview Neural Representations to Demystify the Brain-Visual Alignment

DGX agent

arXiv:2606.25718v1 Announce Type: new Abstract: Zero-shot visual decoding from electroencephalography (EEG) aims to infer visual semantics from non-invasive neural recordings, but remains challenging

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

What happens when Claude Code gets an experiment tracker

DGX agent

At CVPR 2026, Lambda ran a live demo for two and a half days: Claude Code teaching Google's Gemma 4 to play a Tetris-like game. Claude Code started with a Gemma 4 model that couldn't play at all. It p

model-releaseslambda-labs
25 Jun 2026
Model Releases

5.2 could be better with more RL ...

DGX agent

5.2 could be better with more RL ... Deepswe's benchmark results are my own experience. I've used all models, GLM 5.2 ≈ Claude Opus 4.6–4.7. Kimi 2.7 code more like inference optimization. Looking for

model-releasesollama--x
24 Jun 2026
Model Releases

Are We Ready For An Agent-Native Memory System?

DGX agent

arXiv:2606.24775v1 Announce Type: new Abstract: Memory for large language model (LLM) agents has rapidly evolved from simple retrieval-augmented mechanisms into a data management system that supports

model-releasesarxiv-cs-cl
24 Jun 2026
Applications

ArtiTwinSplat: Interactable Digital Twin Reconstruction via Gaussian Splatting from RGB-D videos

DGX agent

arXiv:2606.24628v1 Announce Type: cross Abstract: Deploying robots in unstructured real-world environments needs accurate, interactive models of the objects. Constructing these models at scale remains

applicationsarxiv-cs-cv
24 Jun 2026
Model Releases

Assessing Distribution Shift in Human Activity Recognition for Domain Generalization

DGX agent

arXiv:2606.24781v1 Announce Type: new Abstract: While the field of Human Activity Recognition (HAR) continues to draw interest from researchers and advance in important ways, some key challenges remai

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

At Hugging Face we've been building our own agent that we use via Slack (Moon Bot). Honestly, building your own is quite simple and you'll b…

DGX agent

At Hugging Face we've been building our own agent that we use via Slack (Moon Bot). Honestly, building your own is quite simple and you'll be happy you did: any model you want (self-hosted if needed),

model-releasesclem-delangue--x
24 Jun 2026
Model Releases

AutoSpecNER: A Fine-Grained Named Entity Recognition Dataset for Vehicle Specification Extraction

DGX agent

arXiv:2606.24387v1 Announce Type: new Abstract: Vehicle advertisements contain rich specification information, but automotive NER resources remain limited. We introduce AutoSpecNER, an expert-annotate

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Average Rankings Mask Per-Subject Optimality: A Friedman-Nemenyi Benchmark of EEG Motor-Imagery BCI Decoders

DGX agent

arXiv:2606.24394v1 Announce Type: cross Abstract: Electroencephalography (EEG) is the dominant non-invasive modality for brain-computer interfaces (BCIs), yet reliable decoding of motor imagery is ham

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

BioMedArena: An Open-source Toolkit for Building and Evaluating Biomedical Deep Research Agents

DGX agent

arXiv:2605.06177v2 Announce Type: replace Abstract: Reproducing and comparing deep research agents today is hard: the same backbone evaluated on the same benchmark can report different accuracies acro

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

CanadaFireSat: Toward high-resolution wildfire forecasting with multiple modalities

DGX agent

arXiv:2506.08690v3 Announce Type: replace Abstract: Canada experienced in 2023 one of the most severe wildfire seasons in recent history, causing damage across ecosystems, destroying communities, and

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

CANDLE: Character-level Arabic Noise Deduplication using Lightweight Encoder

DGX agent

arXiv:2606.24758v1 Announce Type: new Abstract: Handling repeated characters in text can be tricky, since they can represent either the correct spelling of a word or informal character elongation ofte

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning

DGX agent

arXiv:2606.24636v1 Announce Type: new Abstract: Cinematographic captioning aims to describe how a video is filmed using professional film-language concepts such as camera movement, shot size, depth of

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

CN-NewsTTS Bench: a target-level automatic benchmark for raw-input Chinese news TTS pronunciation

DGX agent

arXiv:2606.24714v1 Announce Type: new Abstract: Chinese news text contains dense written forms such as scores, hyphenated model names, ranges, unit symbols, percentages, English abbreviations, and mix

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Data Scale, Not Latency, Shapes Cross-Lingual Encoder Transfer in Streaming ASR

DGX agent

arXiv:2606.24169v1 Announce Type: new Abstract: Adapting a streaming speech recognition model to a new language requires choosing between two plausible warm starts: a multilingual (ML) encoder or an E

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

DiffusionBench: On Holistic Evaluation of Diffusion Transformers

DGX agent

arXiv:2606.24888v1 Announce Type: new Abstract: Diffusion transformer (DiT) research on image generation has converged to a single evaluation setup: class-conditional generation on ImageNet. While met

model-releasesarxiv-cs-cv
24 Jun 2026
Tools

GLM 5.2 Fast via Wafer now available on AI Gateway

DGX agent

GLM 5.2 Fast, a faster variant of the GLM language model, is now available through Wafer on Vercel's AI Gateway. This integration allows developers to access the optimized model through Vercel's platf

toolsvercel-blog
24 Jun 2026
Research

Grad Detect: Gradient-Based Hallucination Detection in LLMs

DGX agent

arXiv:2606.24790v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across diverse tasks, yet they remain prone to generating hallucinations. Detec

researcharxiv-cs-ai
24 Jun 2026
Model Releases

LaGO: Latent Action Guidance for Online Reinforcement Learning

DGX agent

arXiv:2606.24669v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong potential for planning and sequential decision-making, but prior work often relies on using them as direc

model-releasesarxiv-cs-ai
24 Jun 2026
Local Ai

Loss Landscape Poisoning: Targeted Extraction of Unseen Training Data from LLMs

DGX agent

arXiv:2606.17110v2 Announce Type: replace-cross Abstract: Large Language Models are increasingly trained on proprietary or sensitive data, from private healthcare and financial records to user convers

local-aiarxiv-cs-lg
24 Jun 2026
Model Releases

Machine Learning and Deep Learning for Exoplanet Detection and Atmospheric Characterization with JWST and the Upcoming Ariel Mission

DGX agent

arXiv:2606.23766v1 Announce Type: cross Abstract: The detection and atmospheric characterization of exoplanets have entered a new data-intensive era driven by the James Webb Space Telescope and the up

model-releasesarxiv-cs-lg
24 Jun 2026
Research

MedPCFM: Improving Medical Point Cloud Completion by Integrating Point Transformers and Flow Matching

DGX agent

arXiv:2606.24433v1 Announce Type: cross Abstract: Medical point cloud completion is important for anatomical reconstruction and downstream clinical workflows, yet generative modeling in this setting r

researcharxiv-cs-ai
24 Jun 2026
Model Releases

Multimedia and Visual Analytics in the Agentic Era

DGX agent

arXiv:2504.06138v3 Announce Type: replace-cross Abstract: Professional users need tools to help them gain actionable insights from large multimedia collections. Foundation models and AI agents have ra

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Neuro-Symbolic Drive: Rule-Grounded Faithful Reasoning for Driving VLAs

DGX agent

arXiv:2606.23938v1 Announce Type: new Abstract: Driving VLA models incorporating Chain-of-Thought (CoT) reasoning are attractive because they leverage pretrained VLM representations and expose interme

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

One Ruler: A Same-Hands Re-Evaluation of Bivariate Causal Direction on Tuebingen, with a Parameter-Free Compression Baseline

DGX agent

arXiv:2606.23767v1 Announce Type: new Abstract: Headline accuracies on the Tuebingen cause-effect pairs are routinely compared across papers even though each is measured under its authors' own protoco

model-releasesarxiv-cs-lg
24 Jun 2026
Tutorials

PointVG-R: Internalizing Geometric Reasoning in MLLMs for Precise Pointing Localization via Visual Chain of Thought

DGX agent

arXiv:2606.24539v1 Announce Type: new Abstract: Pointing-based visual grounding requires models to precisely locate target objects by deciphering complex spatial relationships between the visual scene

tutorialsarxiv-cs-cv
24 Jun 2026
Model Releases

REDI-Match: Rotation-Equivariant Distillation for Efficient and Robust Dense Matching

DGX agent

arXiv:2606.24330v1 Announce Type: new Abstract: Vision Foundation Models (VFMs) have significantly advanced dense feature matching, yet severe in-plane rotation remains a critical challenge. Existing

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Self-Recognition Finetuning can Prevent and Reverse Emergent Misalignment

DGX agent

arXiv:2606.23700v1 Announce Type: cross Abstract: Emergent misalignment (EM) has been linked to the activation of misaligned persona vectors and evil character traits, suggesting that EM operates thro

model-releasesarxiv-cs-ai
24 Jun 2026
Research

Stabilizing Black-Box Prompt Optimization with Textual Regularization and Signal Aggregation

DGX agent

arXiv:2507.09839v2 Announce Type: replace Abstract: An increasing number of NLP applications interact with large language models (LLMs) through black-box APIs, making prompt engineering critical for c

researcharxiv-cs-lg
24 Jun 2026
Model Releases

T2D-Bench: Evidence-Gated Evaluation of LLM Outputs for Type 2 Diabetes Using a Multi-Layer Clinical-Lifestyle Knowledge Graph

DGX agent

arXiv:2606.24145v1 Announce Type: new Abstract: Large language models (LLMs) can produce clinically fluent recommendations for type 2 diabetes while failing to satisfy guideline constraints or explici

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Tri-Efficient Transfer Learning for Point Cloud Videos

DGX agent

arXiv:2606.24175v1 Announce Type: new Abstract: While point cloud foundation models have significantly advanced point cloud video understanding, existing parameter-efficient fine-tuning (PEFT) methods

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

We've provided some updated results on Mistral OCR that make use of the annotation feature for charts. The overall score is ahead of GPT-5.5…

DGX agent

We've provided some updated results on Mistral OCR that make use of the annotation feature for charts. The overall score is ahead of GPT-5.5 and just behind Gemini 3.1 Pro, which is quite impressive f

model-releasesjerry-liu--x
24 Jun 2026
Model Releases

When Retrieval Metrics Mislead: Measuring Policy Signal in Long-Horizon Tool-Use Agents

DGX agent

arXiv:2606.23937v1 Announce Type: cross Abstract: Exact-match retrieval recall is often used as a proxy for whether a retriever supplies useful policy context to a downstream decision model. We test t

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

When Top-1 Fails: Calibrating LoRA Monitors for Masked Diffusion LMs

DGX agent

arXiv:2606.24119v1 Announce Type: cross Abstract: Discrete diffusion language model (DLM) fine-tuning inherits inexpensive diagnostics from denoising-time confidence monitors, but their PEFT-training

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

AI Agents Can Already Autonomously Perform Experimental High Energy Physics

DGX agent

arXiv:2603.20179v3 Announce Type: replace-cross Abstract: Large language model-based AI agents are now able to autonomously execute substantial portions of a high energy physics (HEP) analysis pipelin

model-releasesarxiv-cs-lg
23 Jun 2026
Research

AI-Augmented Thyroid Scintigraphy for Robust Classification of Disease

DGX agent

arXiv:2503.00366v3 Announce Type: replace-cross Abstract: Thyroid scintigraphy is vital for diagnosing thyroid disorders, yet deep learning (DL) models in this domain often struggle with limited, imba

researcharxiv-cs-cv
23 Jun 2026
Model Releases

Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation

DGX agent

arXiv:2606.15127v2 Announce Type: replace Abstract: Reasoning models are increasingly used in settings where the final answer is not the only object of review: educational tools may show students inte

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Concept Alignment Contrast and Long-Short Prompt Memory for Test-Time Adaptation of SAM3 in Medical Image Segmentation

DGX agent

arXiv:2606.22963v1 Announce Type: new Abstract: Concept segmentation models like Segment Anything Model 3 (SAM3) show strong generalization on natural images, yet their performance degrades in medical

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

CoRDE: Concept-Prior Routed Diffusion Experts for Structural Generalization in Robot Manipulation

DGX agent

arXiv:2606.21935v1 Announce Type: new Abstract: Diffusion models excel at capturing multi-modal action distributions in robot imitation learning. However, in multi-task and long-horizon scenarios, mon

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

CVSBench: A Comprehensive Benchmark for Cross-view Spatial Reasoning and Dreaming

DGX agent

arXiv:2606.22476v1 Announce Type: new Abstract: Humans can effortlessly reason about scenes across different viewpoints, yet it remains unclear whether Vision-Language Models (VLMs) possess similar cr

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Data Pruning: Redundant, Problematic, and Interdependent Samples

DGX agent

arXiv:2606.21916v1 Announce Type: new Abstract: The performance of deep learning models is affected by not only data quantity but also data quality. Data pruning is a process by which practitioners ca

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Decodable but Not Faithful: Coupling Natural-Language Rationales to Programmatic Verifiers

DGX agent

arXiv:2606.21678v1 Announce Type: new Abstract: Language models can generate plausible rationales for their predictions, but these explanations may not faithfully represent the model's internal reason

researcharxiv-cs-lg
23 Jun 2026
← Previous
1…516517518519520…1369
Next →