AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Model Releases

Extractable Memorization From First Principles

DGX agent

arXiv:2607.12649v1 Announce Type: cross Abstract: Recent work on extractable memorization in LLMs suffers from two contrasting validity problems. Some studies overstate extraction, e.g., relying on se

model-releasesarxiv-cs-cl
15 Jul 2026
Research

Fast and Accurate Image Restoration and Generation with Rank Enhanced Linear Attention

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2505.16157v2 Announce Type: replace Abstract: Transformer-based models have made remarkable progress in image restoration (IR) tasks. However, the quadratic complexity of self-attention in Trans

researcharxiv-cs-cv
15 Jul 2026
Agents

Fin-Analyst at FinMMEval 2026 Task 3: A Live Hybrid Trading Agent with LLM Specialists and Rule-Based Signals

DGX agent

arXiv:2607.12233v1 Announce Type: cross Abstract: Large language model (LLM) trading agents show promising performance in equity markets, yet remain narrowly focused on US equities with little evidenc

agentsarxiv-cs-ai
15 Jul 2026
Model Releases

GAINS: Gaussian-based Inverse Rendering from Sparse Multi-View Captures

DGX agent

arXiv:2512.09925v2 Announce Type: replace Abstract: Recent advances in Gaussian Splatting-based inverse rendering extend Gaussian primitives with shading parameters and physically grounded light trans

model-releasesarxiv-cs-cv
15 Jul 2026
Research

High-Dimensional Gaussian Mean Estimation under Realizable Contamination

DGX agent

arXiv:2603.16798v2 Announce Type: replace Abstract: We study mean estimation for a Gaussian distribution with identity covariance in R^d under a missing data scheme termed realizable epsilon-contamina

researcharxiv-cs-lg
15 Jul 2026
Model Releases

How I tricked Claude into leaking your deepest, darkest secrets

DGX agent

How I tricked Claude into leaking your deepest, darkest secrets I've been impressed by the way the Claude web_fetch tool is designed to avoid data exfiltration attacks. Ayush Paul found a hole in that

model-releasessimon-willison
15 Jul 2026
Research

Let RGB Be the Language of Vision

DGX agent

arXiv:2607.12450v1 Announce Type: new Abstract: This work introduces a unified formulation for vision models, where diverse forms of visual information beyond natural images, such as masks, depth maps

researcharxiv-cs-cv
15 Jul 2026
Tutorials

QDEvo: A Multi-Objective Quality-Diversity Framework for Automated Heuristic Design

DGX agent

arXiv:2607.11916v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) with evolutionary computation has emerged as a powerful paradigm for automated heuristic design in com

tutorialsarxiv-cs-ai
15 Jul 2026
Applications

ReflectWorld-MM: An Entity-Oriented Multimodal Memory System for Open-Ended Video Streams

DGX agent

arXiv:2607.09759v2 Announce Type: replace-cross Abstract: Building assistants that can continually watch the world, remember what they see, and reason over their accumulated experience is a long-stand

applicationsarxiv-cs-ai
15 Jul 2026
Model Releases

SeqGPT: A Constrained Transformer Agent for the Inverse Designof Multi-Panel Composite Structures

DGX agent

arXiv:2607.11910v1 Announce Type: cross Abstract: Optimizing composite stacking sequences to match continuous targets (e.g., Lamination or Buckling Parameters) with discrete manufacturing constraints

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Cod…

DGX agent

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Code - Claude Fable - @anthropicai culture & product strategy -

model-releasesswyx--x
15 Jul 2026
Local Ai

Toward Production-Ready Federated Learning in Healthcare: Privacy, Orchestration, and Governance in MLOps

DGX agent

arXiv:2607.10467v2 Announce Type: replace-cross Abstract: Healthcare organizations often cannot freely centralize patient data because medical records are sensitive, regulated, and institutionally con

local-aiarxiv-cs-lg
15 Jul 2026
Model Releases

Transforming LLMs into Efficient Cross-Encoders via Knowledge Distillation for RAG Reranking

DGX agent

arXiv:2607.11933v1 Announce Type: new Abstract: Cross-encoders achieve high reranking accuracy in Retrieval-Augmented Generation (RAG) pipelines but impose quadratic inference costs that limit real-ti

model-releasesarxiv-cs-cl
15 Jul 2026
Safety

TrustVLA: Mechanism-Guided Inference-Time Defense Against Vision-Language-Action Backdoors

DGX agent

arXiv:2607.12571v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are deployed through pipelines that end users cannot audit, and a poisoned VLA can behave normally on clean observat

safetyarxiv-cs-ro
15 Jul 2026
Model Releases

UniVR: Thinking in Visual Space for Unified Visual Reasoning

DGX agent

arXiv:2607.12800v1 Announce Type: new Abstract: Learning broad world knowledge directly from raw visual data is a fundamental capability of intelligence. We introduce UniVR, the first investigation in

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

ViHoRec: A Quality-Controlled Vietnamese Hotel Recommendation Dataset and Cold-Start Benchmark

DGX agent

arXiv:2607.12946v1 Announce Type: cross Abstract: Recommender-system research for Vietnamese remains limited by the absence of a public, well-documented hotel interaction resource. Building such a res

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

b10011

DGX agent

server : refactor prompt cache state ownership (#25649) server : clear checkpoints upon prompt clear server : move the prompt state data to the server_prompt_cache Assisted-by: pi:llama.cpp/Qwen3.6-27

model-releasesllama-cpp-releases
14 Jul 2026
Model Releases

Just a reminder that deepseek v3 came out 18 months ago and was considered revolutionary at the time but is basically unusable today There w…

DGX agent

Just a reminder that deepseek v3 came out 18 months ago and was considered revolutionary at the time but is basically unusable today There was a fierce debate at the time about vibe coding and the arg

model-releaseselon-musk--x
14 Jul 2026
Model Releases

By the end of the year we should have: GPT 6 Fable 5.5 Gemini 3.5 Pro Grok 5 Spark 2 Kimi 3 Minimax M3.5 GLM 6 DeepSeek v4.5 Mistral 4 Qwen …

DGX agent

By the end of the year we should have: GPT 6 Fable 5.5 Gemini 3.5 Pro Grok 5 Spark 2 Kimi 3 Minimax M3.5 GLM 6 DeepSeek v4.5 Mistral 4 Qwen 4 MiMo 3 Never in the history of LLMs has the frontier been

model-releasesswyx--x
13 Jul 2026
Model Releases

ChatGPT is available again on WhatsApp in the EEA, part of our work to make AI accessible in the apps people already use every day. Message …

DGX agent

ChatGPT is available again on WhatsApp in the EEA, part of our work to make AI accessible in the apps people already use every day. Message the verified 1-800-CHATGPT contact to ask questions, upload

model-releasesopenai--x
13 Jul 2026
Model Releases

Empowering India’s next generation of innovators with ATL Saathi

DGX agent

ATL Saathi is an initiative aimed at boosting scientific innovation and learning in India through AI-powered tools and programs. Launched in February 2026, the program focuses on accelerating research

model-releasesgoogle-deepmind
13 Jul 2026
Industry

this is huge: best news I've heard all day..nice work, Hugging Face team.

DGX agent

this is huge: best news I've heard all day..nice work, Hugging Face team. Big unlock for open-source AI inference: Hugging Face Transformers models can now run in vLLM at native speed, often matching

industryclem-delangue--x
13 Jul 2026
Model Releases

Hit my Claude and Copilot limits three times in each of the last three days Of course a friendly prompt popped up suggesting I either buy mo…

DGX agent

Hit my Claude and Copilot limits three times in each of the last three days Of course a friendly prompt popped up suggesting I either buy more credits or upgrade my plan Simply switched to Gemini for

model-releasesgary-marcus--x
12 Jul 2026
Industry

Grok 4.5 Review

DGX agent

Grok 4.5 is an AI language model developed by xAI, featuring improvements in reasoning, coding, and multimodal capabilities compared to earlier versions. The model demonstrates enhanced performance ac

industryelon-musk--x
11 Jul 2026
Research

Can We Trust LLM's Logic? Quantifying Uncertainty, Coherence, and Robustness via a Graph-Based Framework

DGX agent

arXiv:2607.08017v1 Announce Type: cross Abstract: Large-Language Models (LLMs) can be prone to flawed and unfaithful reasoning that decoding strategies like Self-Consistency (SC) fail to detect as the

researcharxiv-cs-ai
10 Jul 2026
Research

Classifier Chain-based Pathological Test Recommendation

DGX agent

arXiv:2607.08299v1 Announce Type: new Abstract: Accurate and timely diagnoses are essential for quality patient care. However, delayed recommendation of diagnostic tests and physicians' subjective int

researcharxiv-cs-lg
10 Jul 2026
Model Releases

Cross-seed explainability using Procrustes-conditioned Joint End-to-end Top-K Sparse Autoencoders

DGX agent

arXiv:2607.08499v1 Announce Type: new Abstract: We present a Procrustes-conditioned Joint End-to-end Top-K Sparse Autoencoder (SAE) for extracting cross-seed universal features from independently trai

model-releasesarxiv-cs-cl
10 Jul 2026
Research

CT-CLIP Representations for Multimodal Lung Cancer Survival Prediction

DGX agent

arXiv:2607.08503v1 Announce Type: new Abstract: Accurate prognosis prediction is important for treatment planning in lung cancer, but deep learning-driven survival modelling is often limited by the sc

researcharxiv-cs-cv
10 Jul 2026
Model Releases

Diarization-Guided Qwen-ASR Adaptation for Multilingual Two-Speaker Conversational Speech

DGX agent

arXiv:2607.08208v1 Announce Type: new Abstract: This paper describes our self-designed system for Task 1 of the MLC-SLM 2026 Challenge for multilingual two-speaker conversational speech. The system co

model-releasesarxiv-cs-cl
10 Jul 2026
Agents

From Triggers to Emotions: A CPM-Grounded Appraisal Multi-Agent for Dynamic Emotional Evolution in Persona-Based Dialogue

DGX agent

arXiv:2607.07824v1 Announce Type: cross Abstract: Large Language Models (LLMs) have substantially advanced persona-based dialogue agents for emotion-sensitive role simulation in healthcare, education,

agentsarxiv-cs-ai
10 Jul 2026
Model Releases

Geometry and Gradient-based Partitioning for Panoramic Outdoor Reconstruction

DGX agent

arXiv:2607.08769v1 Announce Type: new Abstract: Scaling 3D Gaussian Splatting (3DGS) to large outdoor scenes is costly in both data acquisition and computation. Adopting panoramic images with equirect

model-releasesarxiv-cs-cv
10 Jul 2026
Industry

Hugging Face’s CEO on why companies are done renting their AI https://techcrunch.com/2026/07/10/hugging-faces-ceo-on-why-companies-are-done-…

DGX agent

Hugging Face's CEO discusses why enterprises are moving away from renting AI models and services from cloud providers, likely advocating for open-source alternatives or on-premise deployment models. T

industryclem-delangue--x
10 Jul 2026
Model Releases

LlamaSeg: Image Segmentation via Autoregressive Mask Generation

DGX agent

arXiv:2505.19422v2 Announce Type: replace Abstract: We present extbf{LlamaSeg}, a visual autoregressive framework that unifies multiple image segmentation tasks via natural language instructions. By r

model-releasesarxiv-cs-cv
10 Jul 2026
Safety

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

DGX agent

arXiv:2607.07903v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit remarkable capabilities but remain highly vulnerable to adversarial prompts and jailbreak attacks. Existing appro

safetyarxiv-cs-ai
10 Jul 2026
Model Releases

MetaHGNIE: Meta-Path Induced Hypergraph Contrastive Learning in Heterogeneous Knowledge Graphs

DGX agent

arXiv:2512.12477v2 Announce Type: replace Abstract: Estimating node importance in heterogeneous knowledge graphs is a fundamental problem underlying recommendation, search, and knowledge decision syst

model-releasesarxiv-cs-ai
10 Jul 2026
Applications

On the Design of Mixture-of-Experts for Dynamic Gaussian Splatting

DGX agent

arXiv:2607.08250v1 Announce Type: new Abstract: Dynamic scene reconstruction remains challenging due to the heterogeneous and spatially varying nature of real-world motion. Although recent 3D Gaussian

applicationsarxiv-cs-cv
10 Jul 2026
Model Releases

PhasorFlow: A Python Library for Unit Circle Based Computing

DGX agent

arXiv:2603.15886v3 Announce Type: replace-cross Abstract: We present PhasorFlow, an open-source Python library for computing on the S^1 unit circle. Inputs are encoded as complex phasors z=e^{iphi} on

model-releasesarxiv-cs-ai
10 Jul 2026
Hardware

Reducing High-Bandwidth Memory Bottlenecks in JAX-Based LLM Training with Host Offloading

DGX agent

Large language model training workloads increasingly encounter GPU memory limits before compute is fully utilized, with high-bandwidth memory capacity becoming the primary scaling bottleneck as model

hardwarenvidia-developer
10 Jul 2026
Hardware

SLORR: Simple and Efficient In-Training Low-Rank Regularization

DGX agent

arXiv:2607.08754v1 Announce Type: cross Abstract: Low-rank factorization is widely used to compress neural networks, but modern models are often not naturally amenable to aggressive factorization with

hardwarearxiv-cs-ai
10 Jul 2026
Model Releases

SwinIFS: Landmark Guided Swin Transformer For Identity Preserving Face Super Resolution

DGX agent

arXiv:2601.01406v2 Announce Type: replace-cross Abstract: Face super-resolution aims to recover high-quality facial images from severely degraded low-resolution inputs, but remains challenging due to

model-releasesarxiv-cs-ai
10 Jul 2026
Applications

The Illusion of Equivalency: Statistical Characterization of Quantization Effects in LLMs

DGX agent

arXiv:2607.08734v1 Announce Type: new Abstract: Post-training quantization is widely used to deploy large language models in resource-constrained settings, yet its evaluation relies almost exclusively

applicationsarxiv-cs-ai
10 Jul 2026
Research

UtterTune: LoRA-Based Target-Language Pronunciation Edit and Control in Multilingual Text-to-Speech

DGX agent

arXiv:2508.09767v3 Announce Type: replace-cross Abstract: We propose UtterTune, a lightweight method for adapting a multilingual text-to-speech (TTS) system built on a large language model (LLM). It i

researcharxiv-cs-cl
10 Jul 2026
Safety

Validating LLMs in social science: Epistemic threats and emerging norms

DGX agent

arXiv:2607.07915v1 Announce Type: cross Abstract: Large language models (LLMs) are reshaping social science methodology. Researchers increasingly prompt language models to generate quantitative measur

safetyarxiv-cs-cl
10 Jul 2026
Model Releases

VSRo-200: A Romanian Visual Speech Recognition Dataset for Studying Supervision and Multimodal Robustness

DGX agent

arXiv:2607.08112v1 Announce Type: new Abstract: We introduce VSRo-200, the first large-scale dataset for visual speech recognition (lip reading) in Romanian, comprising 200 hours of real-world podcast

model-releasesarxiv-cs-cv
10 Jul 2026
Research

What LLM Forecasters Know but Don't Say: Probing Internal Representations for Calibration and Faithfulness

DGX agent

arXiv:2607.08046v1 Announce Type: cross Abstract: Large language models fine-tuned for forecasting can be accurate yet poorly calibrated, and their chain-of-thought (CoT) reasoning may not faithfully

researcharxiv-cs-ai
10 Jul 2026
Research

ZipDepth: Bringing Lightweight Zero-Shot Monocular Depth Anywhere, on Any Device

DGX agent

arXiv:2607.08771v1 Announce Type: new Abstract: Monocular depth estimation has seen remarkable progress through foundation models achieving robust zero-shot generalization, yet their computational dem

researcharxiv-cs-cv
10 Jul 2026
Model Releases

A Gold-Standard Study of What Makes a Lightweight Game-Playing Agent Strong

DGX agent

arXiv:2607.06854v1 Announce Type: cross Abstract: Reinforcement learning agents for imperfect-information card games are only as strong as the opponents they train against, and they are hard to grade,

model-releasesarxiv-cs-ai
9 Jul 2026
Local Ai

A Quiet Failure in Calibrated Virtual Screening: Marginal Conformal Prediction Under-Covers the Minority Class, and a Class-Conditional Fix Recovers It

DGX agent

arXiv:2607.06605v1 Announce Type: new Abstract: Conformal prediction is being adopted in drug discovery to put an honest number on model reliability: pick an error rate alpha, and the method returns p

local-aiarxiv-cs-lg
9 Jul 2026
← Previous
1…587588589590591…1371
Next →