AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlog
89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
Model Releases

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation

DGX agent

arXiv:2606.28998v1 Announce Type: cross Abstract: Large Language Model (LLM) alignment trains an LLM using preference data to produce outputs that better meet established quality standards. While LLM

model-releasesarxiv-cs-ai
30 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

SafePyramid: A Hierarchical Benchmark for In-context Policy Guardrailing

DGX agent

arXiv:2606.29887v1 Announce Type: new Abstract: In real-world applications, guardrails are often expected to identify unsafe user-model interactions according to application-specific safety policies,

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Semantic-Driven Scale and Spatial Selection for Efficient Cross-Modal Alignment in Referring Remote Sensing Image Segmentation

DGX agent

arXiv:2606.30244v1 Announce Type: new Abstract: Referring Remote Sensing Image Segmentation (RRSIS) seeks to localize and segment the target object or region specified by a natural language expression

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Symbolic Mechanistic Data Attribution: Tracing Training Influence to Learned Behavioral Policies

DGX agent

arXiv:2606.29171v1 Announce Type: cross Abstract: While existing data attribution methods can identify which training examples build specific mechanistic circuits, they cannot explain how training dat

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

TF-MoE: Time-Frequency Mixture-of-Experts for Efficient Speech Separation

DGX agent

arXiv:2606.29575v1 Announce Type: cross Abstract: Recent advances in speech separation (SS) have led to compact front-end models with small parameter sizes, yet their high computational cost remains a

model-releasesarxiv-cs-ai
30 Jun 2026
Research

The Fundamental Limits of Valid Transport Map Estimation

DGX agent

arXiv:2606.30574v1 Announce Type: new Abstract: Many modern generative modeling methods, including diffusion models, normalizing flows, and flow matching, estimate transport maps or plans between dist

researcharxiv-cs-lg
30 Jun 2026
Safety

Towards Physical Intuitions for Alignment Dynamics: A Case Study With Randomness Crystallization

DGX agent

arXiv:2606.29933v1 Announce Type: new Abstract: The alignment of language models is typically studied through the lens of capability benchmarks, but the dynamics of how models change during post-train

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

TriageRA-CCF: Source-Side Clinical Confidence and Coverage Signals for Adaptive Rank Budgeting in Medical LLMs

DGX agent

arXiv:2606.29375v1 Announce Type: new Abstract: Medical large language models are commonly adapted with a fixed low-rank budget, even though medical questions differ substantially in confidence, clini

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

UniGP: Taming Diffusion Transformer for Prior-Preserved Unified Generation and Perception

DGX agent

arXiv:2606.30332v1 Announce Type: new Abstract: Recent advances in diffusion models have shown impressive performance in controllable image generation and dense prediction tasks. However, existing app

safetyarxiv-cs-cv
30 Jun 2026
Model Releases

Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare

DGX agent

arXiv:2606.28666v1 Announce Type: cross Abstract: Agent-based AI has enabled the automation of tasks by exposing application tools and resources to large language models (LLMs). However, to improve sc

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

You Only Touch Once: 6-DoF Object Pose Estimation from Single Tactile Contact

DGX agent

arXiv:2606.28899v1 Announce Type: new Abstract: Accurate 6-DoF object pose estimation is fundamental to robotic manipulation, yet vision-based methods often fail under occlusion, poor lighting, and re

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

A Comparison of Fusion Techniques for Multi-Modal Human Activity Recognition on the HARMES Dataset

DGX agent

arXiv:2606.27886v1 Announce Type: new Abstract: Recent advances in Human Activity Recognition (HAR) from wearable sensors have shown that multi-modal deep learning models consistently outperform their

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Accelerating Attention with Basis Decomposition

DGX agent

arXiv:2510.01718v2 Announce Type: replace Abstract: Attention is a core operation in large language models (LLMs). We present BD Attention (BDA), a lossless algorithmic reformulation of attention. BDA

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's t…

DGX agent

Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's the largest AI model theft ever attempted. The same week, the

model-releasesemad-mostaque--x
29 Jun 2026
Model Releases

Applicability of memorization indicators for early spotting of overfitting while recalibrating sEMG-decoders on low sample sizes

DGX agent

arXiv:2606.27855v1 Announce Type: cross Abstract: Deep learning models for surface electromyography (sEMG) can benefit substantially from subject-specific (re-)calibration, since no sufficiently large

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Claude in Microsoft Foundry is now generally available, hosted on Azure. Azure customers get Claude Opus 4.8 and Claude Haiku 4.5, with Azur…

DGX agent

Claude models are now generally available through Microsoft Foundry on Azure infrastructure, providing Azure customers access to Claude Opus 4.8 and Claude Haiku 4.5. This integration allows enterpris

model-releasesboris-cherny--x
29 Jun 2026
Model Releases

Cloud CISO Perspectives: How Google Cloud Security uses AI internally

DGX agent

Welcome to the second Cloud CISO Perspectives for June 2026. Today, we’re discussing how we use AI to chart a path to autonomous software development lifecycle security.As with all Cloud CISO Perspect

model-releasesgoogle-cloud-ai
29 Jun 2026
Model Releases

Contagion Networks: Evaluator Preference Propagation in Multi-Agent LLM Systems

DGX agent

arXiv:2606.20493v2 Announce Type: replace-cross Abstract: When large language models serve as evaluators in multi-agent systems, their strategy preferences -- whether induced by explicit prompts or by

model-releasesarxiv-cs-ai
29 Jun 2026
Safety

Deployment-Side Adaptiveness in Multi-Horizon Volatility Forecasting

DGX agent

arXiv:2606.27688v1 Announce Type: cross Abstract: In financial forecasting, predictive performance depends not only on which model is trained, but also on how the trained model is deployed. We study t

safetyarxiv-cs-ai
29 Jun 2026
Model Releases

Estimation--Prediction Tradeoff in Causal Probabilistic Temporal Graphs

DGX agent

arXiv:2606.28225v1 Announce Type: new Abstract: Temporal link prediction is usually evaluated by predictive performance on unseen edges, but in probabilistic temporal graphs this criterion can conflat

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning

DGX agent

arXiv:2603.09731v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) are increasingly considered as a foundation for embodied agents, yet it remains unclear whether they

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs

DGX agent

arXiv:2606.27378v1 Announce Type: new Abstract: We introduce an axiomatic evaluation framework for latent thought representations in LLMs, comprising metrics that are independent of downstream benchma

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

From Detection to Action: Using LLM Agents for Fault-Tolerant Control

DGX agent

arXiv:2606.28011v1 Announce Type: cross Abstract: We propose an agentic Large Language Model (LLM) framework for active Fault-Tolerant Control (FTC) that transforms fault detection outputs into constr

model-releasesarxiv-cs-lg
29 Jun 2026
Agents

GraphPilot: Grounded Scene Graph Conditioning for Language-Based Autonomous Driving

DGX agent

arXiv:2511.11266v4 Announce Type: replace Abstract: Vision-language models have recently emerged as promising planners for autonomous driving, where success hinges on topology-aware reasoning over spa

agentsarxiv-cs-cv
29 Jun 2026
Model Releases

Learning to Evict from Key-Value Cache

DGX agent

arXiv:2602.10238v2 Announce Type: replace Abstract: The growing size of Large Language Models (LLMs) makes efficient inference challenging, primarily due to the memory demands of the autoregressive Ke

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

OrthoTryOn: Geometric Orthogonalization for Conflict-Free Unified Fashion Generation

DGX agent

arXiv:2606.27880v1 Announce Type: new Abstract: Unified fashion generation integrates tasks like virtual try-on and garment reconstruction into a single model to reduce task-specific adaptation costs.

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Output-Space Allocation Costs for Calibration-Guided LLM Compression: An Empirical Study

DGX agent

arXiv:2606.27785v1 Announce Type: cross Abstract: Training-free compression methods for large language models (LLMs) often use calibration data to guide compression decisions. ROCKET, a recent method

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Position Bias Correction is Insufficient for One-Pass Attention Sorting

DGX agent

arXiv:2606.27793v1 Announce Type: cross Abstract: Long-context language models suffer from position bias, where information in middle positions is underutilized. Attention Sorting addresses this by it

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Retaining by Doing: The Role of On-Policy Data in Mitigating Forgetting

DGX agent

arXiv:2510.18874v3 Announce Type: replace-cross Abstract: Adapting language models (LMs) to new tasks via post-training carries the risk of degrading existing capabilities -- a phenomenon classically

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

Triadic Werewolf: A Jester Role for Multi-Hop Theory of Mind in LLMs

DGX agent

arXiv:2606.27909v1 Announce Type: cross Abstract: Theory-of-mind evaluations of large language models typically use dyadic social-deduction games, where every observable cue points to a single hidden

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search

DGX agent

arXiv:2606.27669v1 Announce Type: new Abstract: Search agents powered by large language models (LLMs) are increasingly used to solve complex information-seeking tasks, requiring multi-step retrieval a

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

ZooClaw-FashionSigLIP2: Distilled Fine-tuning for Robust Fashion Retrieval

DGX agent

arXiv:2606.27708v1 Announce Type: new Abstract: Adapting a foundation vision-language encoder to a specialized retrieval task creates a fundamental tradeoff: gains on the target distribution come at t

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

China’s Z.ai claims it can match Mythos on cybersecurity

DGX agent

China's Zhipu AI (Z.ai) released its open-weight GLM-5.2, and some researchers have claimed that it matches Mythos in certain bug-finding and cybersecurity scenarios. While GLM lags behind models from

model-releasesthe-verge-ai
28 Jun 2026
Model Releases

Elon Musk turns 55 today. Here are 55 milestones. Age 54: world's first trillionaire Age 54: takes SpaceX public Age 54: SpaceX acquires xAI…

DGX agent

Elon Musk turns 55 today. Here are 55 milestones. Age 54: world's first trillionaire Age 54: takes SpaceX public Age 54: SpaceX acquires xAI Age 54: releases Grok 4 Age 53: launches Robotaxi service A

model-releaseselon-musk--x
28 Jun 2026
Model Releases

https://huggingface.co/collections/deepseek-ai/deepspec

DGX agent

https://huggingface.co/collections/deepseek-ai/deepspec Good guy DeepSeek gives us accelerated models The most interesting one here is Gemma4-12B, I presume vision included. Might be the best local mo

model-releasesclem-delangue--x
28 Jun 2026
Model Releases

[AINews] OpenAI GPT-5.6 Sol / Terra / Luna — restricted to trusted partners

DGX agent

OpenAI has released GPT-5.6 with three variants (Sol, Terra, Luna) in a restricted beta program limited to trusted partners, according to reporting from Latent Space. The restricted access model sugge

model-releaseslatent-space
27 Jun 2026
Model Releases

A probabilistic framework for online test-time adaptation

DGX agent

arXiv:2606.26457v1 Announce Type: cross Abstract: This paper presents a probabilistic framework for online test-time adaptation problems. In them, a model is trained on labeled data but must adapt to

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Adaptive Evaluation of Out-of-Band Defenses Against Prompt Injection in LLM Agents

DGX agent

arXiv:2606.26479v1 Announce Type: cross Abstract: Recent work (2024 to 2026) has converged on a strategy for defending tool-using LLM agents against indirect prompt injection: rather than training the

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Comparing BERT Sentence-Pair Classification and Few-Shot LLM Prompting for Detecting Threat and Solution Framing in German Climate News

DGX agent

arXiv:2606.26489v1 Announce Type: new Abstract: News media play a central role in shaping public perceptions of climate change, and whether coverage emphasizes threats or solutions has measurable effe

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Confidence-Aware Tool Orchestration for Robust Video Understanding

DGX agent

arXiv:2606.26904v1 Announce Type: cross Abstract: Video reasoning language models implicitly assume that every input frame is equally reliable. This leads to what we term the Blind Trust Problem: unde

model-releasesarxiv-cs-ai
26 Jun 2026
Applications

Discovering Millions of Interpretable Features with Sparse Autoencoders

DGX agent

arXiv:2606.26620v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have emerged as a powerful tool for decomposing superposed language model representations into sparse and interpretable fea

applicationsarxiv-cs-ai
26 Jun 2026
Model Releases

Does AI Reviewer See the Full Picture? Attacking and Defending Multimodal Peer Review

DGX agent

arXiv:2606.12716v2 Announce Type: replace Abstract: The integration of Large Language Models (LLMs) and Multimodal LLMs (MLLMs) into scientific peer-review workflows introduces novel and significant r

model-releasesarxiv-cs-cl
26 Jun 2026
Agents

EvoEmbedding: Evolvable Representations for Long-Context Retrieval and Agentic Memory

DGX agent

arXiv:2606.21649v2 Announce Type: replace Abstract: Existing embedding models are inherently static: they encode text segments in isolation, ignoring their surrounding context and temporal order. This

agentsarxiv-cs-cl
26 Jun 2026
Model Releases

GAVEL: Grounded Caption Error Verification and Localization

DGX agent

arXiv:2606.26923v1 Announce Type: new Abstract: Vision-language models (VLMs) often produce hallucinated or inconsistent outputs, where text and images are not properly aligned. Addressing this issue

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Have been taking different local open-weight LLMs for a test drive in different harnesses (Qwen-Code, Codex, Claude Code). 30B Mixture-of-Ex…

DGX agent

Have been taking different local open-weight LLMs for a test drive in different harnesses (Qwen-Code, Codex, Claude Code). 30B Mixture-of-Expert models are kind of a nice sweet spot and can solve chal

model-releasessebastian-raschka--x
26 Jun 2026
Model Releases

LA4VLA: Learning to Act without Seeing via Language-Action Pretraining

DGX agent

arXiv:2606.27295v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are commonly pretrained on robot demonstrations by jointly mapping visual observations and language instructions to

model-releasesarxiv-cs-ro
26 Jun 2026
Model Releases

Latent Diffusion Posterior Sampling with Surrogate Likelihood Guidance for PDE Inverse Problems

DGX agent

arXiv:2606.26592v1 Announce Type: cross Abstract: We propose latent-space diffusion posterior sampling (L-DPS), an approximate Bayesian framework for high-dimensional inverse problems governed by part

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

LMs as Task-Specific Knowledge Bases: An Interpretability Analysis

DGX agent

arXiv:2606.27237v1 Announce Type: new Abstract: Language models (LMs) capture large amounts of factual knowledge applicable to a wide range of tasks, motivating the view of their parameters as a knowl

model-releasesarxiv-cs-cl
26 Jun 2026
← Previous
1…438439440441442…1338
Next →