AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
91,566Total entries
1Added by human
91,565Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,232 results
Safety

PriFT: Prior-Support Guided Supervised Fine-Tuning

DGX agent

arXiv:2606.09396v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is an efficient approach for downstream task adaptation and often serves as the initialization stage for reinforcement le

safetyarxiv-cs-lg
9 Jun 2026
Model Releases

Programmable Silicon Retina on Pixel Processor Array

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.08370v1 Announce Type: cross Abstract: Standard dynamic vision sensors approximate retinal processing by detecting temporal contrast changes, offering high speed and high dynamic range. In

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Quantum feature-map learning with reduced resource overhead

DGX agent

arXiv:2510.03389v2 Announce Type: replace-cross Abstract: Current quantum computers require algorithms that use limited resources economically. In quantum machine learning, success hinges on quantum f

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Real-Time Industrial Defect Detection on Edge Hardware Using Fine-Tuned YOLOv8: A Systematic Benchmark on the NEU Surface Defect Database and MVTec AD with Automotive & Battery Manufacturing Extensions

DGX agent

arXiv:2606.07659v1 Announce Type: new Abstract: Automated surface defect detection is critical for ensuring rigorous quality control in high-speed manufacturing environments. While deep learning model

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Robust-U1: Can MLLMs Self-Recover Corrupted Visual Content for Robust Understanding?

DGX agent

arXiv:2606.08063v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in visual understanding, yet their performance degrades significantly un

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Rosetta Memory: Adaptive Memory for Cross-LLM Agents

DGX agent

arXiv:2606.07711v1 Announce Type: cross Abstract: Memory is the key component for transforming a stateless LLM into a persistent, evolving agent through experience accumulation, long-horizon planning,

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SafeRun: Enabling Determinism in LLM Planning for Running

DGX agent

arXiv:2606.09027v1 Announce Type: cross Abstract: Large Language Models enable flexible natural-language planning but remain unreliable in determinism-critical domains due to their probabilistic natur

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SegmentAnyTreeV2: Scaling Transformer-Based Tree Instance Segmentation Across Sensors, Platforms, and Forests

DGX agent

arXiv:2606.08206v1 Announce Type: new Abstract: We present SegmentAnyTreeV2, a sensor- and platform-agnostic framework for semantic and instance segmentation of forest point clouds. The model combines

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Shift-Dependent Asymmetry: Orthogonal Inverse Low-Rank Adaptation for Federated Medical Segmentation

DGX agent

arXiv:2606.08687v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) enables efficient federated fine-tuning of segmentation foundation models for medical imaging. However, most federated LoRA m

model-releasesarxiv-cs-cv
9 Jun 2026
Research

SOMA: From Surface Observations to Muscle Anatomy

DGX agent

arXiv:2606.09246v1 Announce Type: new Abstract: With the growing demand for realistic virtual humans, parametric body models have become a cornerstone of modern medicine, sports, and entertainment app

researcharxiv-cs-cv
9 Jun 2026
Model Releases

SSAFE: Simple and Strong AI-Generated Image Detection via Frozen Vision Encoders

DGX agent

arXiv:2606.08634v1 Announce Type: new Abstract: The rapid advancement of generative models has blurred the boundary between synthetic and real imagery, creating an urgent need for reliable deepfake de

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Stabilizing On-Policy Distillation for MLLM Reasoning with Global Normalization

DGX agent

arXiv:2606.09091v1 Announce Type: cross Abstract: On-policy distillation (OPD) has recently emerged as an important post-training paradigm. By using a stronger teacher model to provide dense, fine-gra

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Structured Neuron Pruning in Deep Neural Networks Using Multi-Armed Bandits

DGX agent

arXiv:2606.07615v1 Announce Type: cross Abstract: Deep neural networks often contain redundant hidden units. Removing individual weights can reduce parameter count, but unstructured sparsity is not al

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Summarization is Not Dead Yet

DGX agent

arXiv:2606.08000v1 Announce Type: cross Abstract: The progress of large language models (LLMs) has fueled claims that model-generated summaries rival or even surpass human-written references, raising

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

The Sample Complexity of Parameter-Free Stochastic Convex Optimization

DGX agent

arXiv:2506.11336v2 Announce Type: replace Abstract: We study the sample complexity of stochastic convex optimization when problem parameters such as the distance to optimality and the Lipschitz consta

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Token Sample Complexity of Attention

DGX agent

arXiv:2512.10656v3 Announce Type: replace Abstract: As context windows in large language models continue to expand, it is essential to characterize how attention behaves at extreme sequence lengths. W

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Training-Free Generalized Few-Shot Segmentation through Open-Vocabulary Semantic Arbitration

DGX agent

arXiv:2606.09474v1 Announce Type: new Abstract: Generalized Few-Shot Semantic Segmentation (GFSS) has traditionally been approached as a representation-learning problem, requiring task-specific adapta

model-releasesarxiv-cs-cv
9 Jun 2026
Hardware

TUDSR: Twice Upsampling-Diffusion for Higher Super-Resolution

DGX agent

arXiv:2606.09608v1 Announce Type: new Abstract: Diffusion-based generative models have achieved remarkable success in real-world image super-resolution (SR). With tiled diffusion techniques, these mod

hardwarearxiv-cs-cv
9 Jun 2026
Tutorials

Weak-Driven Learning: How Weak Agents make Strong Agents Stronger

DGX agent

arXiv:2602.08222v2 Announce Type: replace Abstract: As post-training optimization becomes central to improving large language models, we observe a persistent saturation bottleneck: once models grow hi

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

What Codex unlocks for Notion

DGX agent

OpenAI's Codex model enables Notion to add AI-powered capabilities to its workspace platform, allowing users to automate tasks and generate content through natural language commands. This integration

model-releasesopenai
9 Jun 2026
Safety

AdaGRPO: A Capability-Aware Adaptive Enhancement for Flow-based GRPO

DGX agent

arXiv:2606.06828v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has demonstrated remarkable success in aligning text-to-image (T2I) flow models with human preferences. Howeve

safetyarxiv-cs-cv
8 Jun 2026
Model Releases

Aumann-SHAP: The Geometry of Counterfactual Interaction Explanations in Machine Learning

DGX agent

arXiv:2603.14014v2 Announce Type: replace Abstract: We introduce Aumann-SHAP, an interaction-aware framework that decomposes counterfactual transitions by restricting the model to a local hypercube co

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Building super fast experiences with Gemma just got easier. Gemma 4 MTP is now officially merged into llama.cpp. Developers can now pair MTP…

DGX agent

Gemma 4 MTP (Multi-Token Prediction) has been officially integrated into llama.cpp, enabling developers to build faster AI experiences by using the model with this inference framework. This merge allo

model-releasesgeorgi-gerganov--x
8 Jun 2026
Research

CoMetaPNS: Continually Meta-learning Personalized Neural Surrogates for Cardiac Electrophysiology Simulations

DGX agent

arXiv:2606.07488v1 Announce Type: new Abstract: Personalized virtual heart simulations face challenges in model personalization and computational cost. While neural surrogates offer state-of-the-art s

researcharxiv-cs-lg
8 Jun 2026
Model Releases

Compute-Optimal Network Design for Echocardiography Myocardial Segmentation and Perfusion Quantification using Neural Scaling Laws

DGX agent

arXiv:2606.06725v1 Announce Type: cross Abstract: Myocardial perfusion quantification using contrast-enhanced ultrasound offers a bedside non-ionizing alternative to nuclear imaging modalities. Howeve

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

CoQuIR: A Comprehensive Benchmark for Code Quality-Aware Information Retrieval

DGX agent

arXiv:2506.11066v3 Announce Type: replace-cross Abstract: Code retrieval is essential in modern software development, as it boosts code reuse and accelerates debugging. However, current benchmarks pri

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

DaX: Learning General Pathology Representations Across Scales

DGX agent

arXiv:2606.06983v1 Announce Type: cross Abstract: Computational pathology requires visual representations that transfer across diverse clinical endpoints and remain robust to variation in magnificatio

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Declarative Skills for AI Agents in Knowledge-Grounded Tool-Use Workflows

DGX agent

arXiv:2606.06923v1 Announce Type: new Abstract: We study orchestration mechanisms for tool-using AI agents in realistic customer-service workflows over an unstructured knowledge base. We argue that de

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

From Vision to Text: A Compact Multimodal Approach for Robust, Cross-Domain Presentation Attack Detection on ID Cards

DGX agent

arXiv:2606.06966v1 Announce Type: new Abstract: Cross-domain shifts challenge Presentation Attack Detection (PAD) on ID Cards, given the restricted data available due to privacy concerns. This work pr

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

Gemma 4 Chat Template now has preserve thinking

DGX agent

Google added an empty thinking token to the Gemma 4 chat template, which stabilizes model output by suppressing 'ghost' thought channels that may appear even when thinking is deactivated. This update

model-releasesr-localllama
8 Jun 2026
Model Releases

Learning Perspectivist Social Meaning via Demographic-Conditioned Fusion Embeddings

DGX agent

arXiv:2606.07123v1 Announce Type: new Abstract: Social meaning in language is inherently perspectival, varying across annotator backgrounds, demographics, and ideological positions. However, most NLP

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

LLM-Guided Evolution for Medical Decision Pipelines

DGX agent

arXiv:2606.07342v1 Announce Type: new Abstract: Adapting large language models (LLMs) to clinical workflows often requires costly fine-tuning or manual prompt and pipeline engineering. We study LLM-gu

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

MacArena: Benchmarking Computer Use Agents on an Online macOS Environment

DGX agent

arXiv:2606.06560v1 Announce Type: cross Abstract: Computer-use agents (CUAs) operate graphical user interfaces (GUIs) through vision and control primitives, and their capabilities have advanced rapidl

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

NotebookLM’s Gemini 3.5 upgrade adds a cloud computer and help finding sources

DGX agent

Google is rolling out 'across the board' updates to NotebookLM. The AI-powered note-taking app now uses Google's upgraded Gemini 3.5 model, which will allow it to respond with 'more accurate and relia

model-releasesthe-verge-ai
8 Jun 2026
Model Releases

On the Geometry of On-Policy Distillation

DGX agent

arXiv:2606.07082v1 Announce Type: cross Abstract: On-policy distillation (OPD) is increasingly used to improve large language model reasoning, but its training dynamics remain poorly understood. We ch

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Phun-Bench: Evaluating LLMs on Phonological Understanding in Chinese

DGX agent

arXiv:2606.07300v1 Announce Type: new Abstract: Language is a vehicle for thought, intricately tied to sounds, symbols, and meaning. However, most large language model (LLM) research focuses on meanin

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

Pipeline parallelism in llama.cpp may be wasting your VRAM

DGX agent

Pipeline parallelism in llama.cpp distributes model layers across multiple GPUs, with each GPU holding a contiguous slice of layers . However, the Reddit post likely discusses inefficiencies in how pi

model-releasesr-localllama
8 Jun 2026
Safety

RASFT: Rollout-Adaptive Supervised Fine-Tuning for Reasoning

DGX agent

arXiv:2606.07006v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is a prevailing method for adapting large language models to reasoning tasks by imitating offline expert demonstrations,

safetyarxiv-cs-cl
8 Jun 2026
Research

Re-Centering Humans in LLM Personalization

DGX agent

arXiv:2606.06614v1 Announce Type: cross Abstract: Despite growing interest, most evaluations of large language models' (LLMs') personalization abilities have relied on synthetic data. It remains uncle

researcharxiv-cs-ai
8 Jun 2026
Model Releases

ScenicRules: An Autonomous Driving Benchmark with Multi-Objective Specifications and Abstract Scenarios

DGX agent

arXiv:2602.16073v2 Announce Type: replace-cross Abstract: Developing autonomous driving systems for complex traffic environments requires balancing multiple objectives, such as avoiding collisions, ob

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

SEAM: Shortcut-Aware Real-Time Detection of Scripted vs. Spontaneous Speech for Interview Guardrails

DGX agent

arXiv:2606.06837v1 Announce Type: cross Abstract: Scripted vs spontaneous speech detection is appealing for interview guardrails, but benchmark performance can be inflated by shortcuts tied to corpus

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Siri AI at WWDC 2026

DGX agent

Given how badly burned anyone who took Apple's 2024 WWDC Apple Intelligence announcements at face value was, I'm holding to a strict 'I'll believe it when I see it' policy for everything they announce

model-releasessimon-willison
8 Jun 2026
Model Releases

So excited to be opening up OpenEnv to the whole community. It will now be owned by @huggingface , Meta-PyTorch, @reflection_ai , @UnslothAI…

DGX agent

So excited to be opening up OpenEnv to the whole community. It will now be owned by @huggingface , Meta-PyTorch, @reflection_ai , @UnslothAI , @modal, @PrimeIntellect , @NVIDIAAI , @mercor_ai , and @f

model-releasesclem-delangue--x
8 Jun 2026
Model Releases

Sparse Subspace-to-Expert Sharing for Task-Agnostic Continual Learning

DGX agent

arXiv:2606.07500v1 Announce Type: cross Abstract: Continual learning in Large Language Models (LLMs) is hindered by the plasticity-stability dilemma, where acquiring new capabilities often leads to ca

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

Stable Reasoning, Unstable Responses: Mitigating LLM Deception via Stability Asymmetry

DGX agent

arXiv:2603.26846v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) expand in capability and application scope, their trustworthiness becomes critical. A vital risk is intrinsic

safetyarxiv-cs-ai
8 Jun 2026
Local Ai

The Masked Advantage: Uncovering Local-Language Access to Cultural Knowledge in LLMs

DGX agent

arXiv:2606.07422v1 Announce Type: cross Abstract: Large language models are increasingly used to answer culturally grounded questions across languages, yet it remains unclear whether local cultural kn

local-aiarxiv-cs-ai
8 Jun 2026
Research

The Necessity of Setting Temperature in LLM-as-a-Judge

DGX agent

arXiv:2603.28304v2 Announce Type: replace Abstract: Using large language models (LLMs) as judges for evaluating model outputs has emerged as an important paradigm for automated evaluation. However, th

researcharxiv-cs-cl
8 Jun 2026
Model Releases

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment

DGX agent

arXiv:2606.06667v1 Announce Type: new Abstract: The mechanisms behind LLMs' broad over-generalization beyond training examples remain unclear. Emergent misalignment (EM) offers a striking case study:

model-releasesarxiv-cs-cl
8 Jun 2026
← Previous
1…525526527528529…1380
Next →