AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,504 results
10 Aug 2026

Confidence Estimation for Financial Vision-Language Models in Chart and Document Understanding

SafetyDGX agent

arXiv:2608.06532v1 Announce Type: new Abstract: LVLMs are increasingly used to read financial charts, tables, and documents, where a single misread figure can move a decision and the most authoritativ

Confirming Our Biases? Evaluating the Capabilities, Risks, and Societal Impact of Large Language Models

ResearchDGX agent

arXiv:2608.06977v1 Announce Type: new Abstract: It is well established that large language models (LLMs) are sensitive to prompt framing, reflecting patterns in their training data or prior prompts. I

Fast and Accurate: An Adaptive VLA Inference Framework through Environment-aware Model Selection

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.06434v1 Announce Type: cross Abstract: Embodied intelligence demands both long-horizon reasoning and real-time closed-loop responsiveness. Recent dual-system Vision-Language-Action (VLA) ar

Foundation Models Adaptation for Multi-View Multi-modal Cardiac MRI Segmentation and Direct Ejection Fraction Estimation

ResearchDGX agent

arXiv:2608.07291v1 Announce Type: new Abstract: Foundation models have shown strong transferability in cardiac MRI (CMR), but their effectiveness for heterogeneous multi-view and multi-sequence CMR an

Measuring the Cross-Lingual Comprehension Gap: How the language of the evidence shapes what language models understand

ResearchDGX agent

arXiv:2608.06506v1 Announce Type: new Abstract: Language models are often evaluated as though capabilities demonstrated in English remain equally available when the same content is presented in other

Recovering Lesion Parameters from Aphasic Picture Naming Error Profiles in Large Language Models

ResearchDGX agent

arXiv:2608.06429v1 Announce Type: new Abstract: Interpretability methods for large language models (LLMs) describe internal state but do not directly test whether that state is causally sufficient to

9 Aug 2026

the term “RLM” (recursive language model) got a lot of buzz this week, but this idea is not new! @a1zhang wrote the og RLM paper 10 months a…

AgentsDGX agent

the term “RLM” (recursive language model) got a lot of buzz this week, but this idea is not new! @a1zhang wrote the og RLM paper 10 months ago! thats like 5 agent-years! would highly recommend followi

8 Aug 2026

Kimi K3 (Unsloth) IQ2-XXS from 711GB down to 478GB!!! Only Multi-language was removed to trim the size

Model ReleasesDGX agent

Firstly a big thanks to the poster 'hellohazine', he basically only removed the multi-lingual fat of the model and just kept the English language intact. It is the exact model, and the rest of the mod

7 Aug 2026

A note on conditional PAC-efficient reasoning in large language model routing

ResearchDGX agent

arXiv:2512.03057v2 Announce Type: replace-cross Abstract: We study distribution-free risk control for model routing, motivated by large language model reasoning. We formalize pointwise conditional eff

Bayesian Expected Uncertainty Reduction (B-EUR) Model: A Computational Account of What Makes Design Options Worth Trying

ResearchDGX agent

arXiv:2608.05642v1 Announce Type: new Abstract: This paper proposes the Bayesian Expected Uncertainty Reduction (B-EUR) model, which formalizes the value of trying a candidate design action as its exp

CPC-CMS: Cognitive Pairwise Comparison Classification Model Selection Framework for Document-level Sentiment Analysis

ResearchDGX agent

arXiv:2507.14022v2 Announce Type: replace Abstract: This study proposes the Cognitive Pairwise Comparison Classification Model Selection (CPC-CMS) framework for document-level sentiment analysis. The

CREBench: Evaluating Large Language Models in Cryptographic Binary Reverse Engineering

Model ReleasesDGX agent

arXiv:2604.03750v2 Announce Type: replace-cross Abstract: Reverse engineering (RE) is central to software security, particularly for cryptographic programs that handle sensitive data and are highly pr

Faster and Better Alignment for Flow Matching Models via Step-aware Advantages

SafetyDGX agent

arXiv:2602.01591v2 Announce Type: replace Abstract: Recent advances in flow matching models, particularly with reinforcement learning (RL), have significantly enhanced human preference alignment in fe

Grounded Well-Condition Anomaly Detection on the Volve Field: Constructed Labels, a Baseline, and a Dual-Head Model

Model ReleasesDGX agent

arXiv:2608.05685v1 Announce Type: new Abstract: Most public benchmarks for machine-condition monitoring come from test rigs, where faults are induced on purpose and every event is known. Real producti

Marginal Matching Does Not License Factorized Sampling: Auditing Conditional Style Leakage in Factorized Generative Models

ApplicationsDGX agent

arXiv:2608.05243v1 Announce Type: cross Abstract: Factorized generative models commonly regularize a latent style variable z_s by matching its marginal distribution to a fixed Gaussian prior and inter

Persona-Pruner: Sculpting Lightweight Models for Role-Playing

ApplicationsDGX agent

arXiv:2606.14695v2 Announce Type: replace-cross Abstract: Language Models (LMs) have shown remarkable potential as role-playing chatbots, delivering consistent, stylized interactions when given a spec

Quantum-Structured World Models (QSWMs) for Predictive Latent Dynamics

TutorialsDGX agent

arXiv:2608.05371v1 Announce Type: new Abstract: World models learn latent states that summarize interaction histories, evolve over time, and support prediction, simulation, or planning. Most existing

Robust-WAM: Bridging Generative Pretraining and Semantic Foresight in World-Action Models

SafetyDGX agent

arXiv:2608.05903v1 Announce Type: new Abstract: Mainstream World-Action Models (WAMs) adapt pretrained video generation models (VGMs) for robot control, transferring their learned dynamics prior for a

SkillTFM: Gated Skill Evolution for Training-Free Adaptation of Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2608.06137v1 Announce Type: new Abstract: Tabular data are ubiquitous in real-world applications and are crucial for data-driven prediction and decision-making across science, industry, finance,

Text-Guided Refinement of Multi-sequence Glioma Subregion Segmentation with a Vision-Language Foundation Model

Model ReleasesDGX agent

arXiv:2608.05389v1 Announce Type: new Abstract: Background: Accurate glioma subregion delineation is important for radiotherapy planning and longitudinal monitoring, but manual contour correction is t

The Ignition Index: Measuring Global Workspace Dynamics in Language Models

Model ReleasesDGX agent

arXiv:2608.05160v1 Announce Type: new Abstract: We introduce the Ignition Index (I), a validated scalar metric that operationalizes Global Workspace Theory's (GWT) all-or-none ignition prediction in t

World-to-Wrist: Task-Conditioned Future Wrist Modeling for Fine-Grained Robot Manipulation

Local AiDGX agent

arXiv:2608.05369v1 Announce Type: cross Abstract: Vision-language-action (VLA) models often treat main-view and wrist-view observations as parallel visual inputs, overlooking their distinct roles in r

6 Aug 2026

Best open-source harnesses for combining cloud and local AI model orchestration?

Local AiDGX agent

Looking for best current solutions for combining cloud models and local models seamlessly inside a harness' orchestration Edit: Right now, we don't have harnesses (that I'm aware of) that are blending

Chained Recursive Language Models for Multi-Iteration Reasoning

ResearchDGX agent

arXiv:2608.05124v1 Announce Type: cross Abstract: Long context reasoning in large language models (LLMs) is usually constrained by the fact that a single inference trajectory has to simultaneously exp

EdgeLM: Edge Demonstrations for Language Models' Table Understanding

ResearchDGX agent

arXiv:2608.04390v1 Announce Type: new Abstract: Large language models (LLMs) perform table-centric prediction through in-context learning, making demonstration selection critical to performance. Exist

Filing: DeepSeek has invested ~$20.8M in Unitree Robotics' Shanghai IPO and agreed to jointly develop AI models for humanoid machines (Eduardo Baptista/Reuters)

Model ReleasesDGX agent

Eduardo Baptista / Reuters: Filing: DeepSeek has invested ~20.8M in Unitree Robotics' Shanghai IPO and agreed to jointly develop AI models for humanoid machines — Chinese artificial intelligence start

It is past time to take AI & security seriously at the individual level as well. If its not the current OpenAI and Anthropic models doing it…

ApplicationsDGX agent

It is past time to take AI & security seriously at the individual level as well. If its not the current OpenAI and Anthropic models doing it, then the coming open weights models will when they catch u

Large Language Models for Low-Resource Languages: A Conceptual Framework for an Electronic Explanatory Dictionary of the Tajik Language

Model ReleasesDGX agent

arXiv:2608.04186v1 Announce Type: new Abstract: This paper presents a conceptual framework for developing an electronic explanatory dictionary of the Tajik language using large language models (LLMs).

MERaLiON-GR: Speech Gender Recognition Model for English and SEA Languages

Model ReleasesDGX agent

arXiv:2608.04433v1 Announce Type: cross Abstract: We present MERaLiON-GR, a speech gender recognition system that performs binary classification (female / male) on English and Southeast Asian (SEA) la

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight

Model ReleasesDGX agent

arXiv:2608.04657v1 Announce Type: new Abstract: World action models (WAMs) built on video generation backbones are a rising recipe for robot learning, yet remain confined to tabletop manipulation. Mob

MultiPathFormer: Towards a Foundation Model for Multipath Wireless Propagation

Local AiDGX agent

arXiv:2608.05076v1 Announce Type: cross Abstract: Recent advances in machine learning have enabled training of wireless foundation models, which aim to support tasks such as channel estimation, beam p

OpenAI updates the default model for free users to GPT-5.6 Luna, adds unlimited text chats for free users, rolls out an improved GPT-5.6 Sol version, and more (Herb Scribner/Axios)

Model ReleasesDGX agent

Herb Scribner / Axios: OpenAI updates the default model for free users to GPT-5.6 Luna, adds unlimited text chats for free users, rolls out an improved GPT-5.6 Sol version, and more — OpenAI on Thursd

Overcoming Statistical Bias in Action-Controllable World Models

SafetyDGX agent

arXiv:2608.04653v1 Announce Type: new Abstract: Action-conditioned world models aim to predict how visual environments evolve under an agent's actions. Yet future frames are often highly predictable f

Persistent Object Narratives for Token-Efficient Video Language Models

Model ReleasesDGX agent

arXiv:2608.04866v1 Announce Type: new Abstract: Video large language models (Video-LLMs) have made strong progress in open-ended video understanding. However, their visual interfaces remain token-inte

Personalized Federated Sparse Adaptation of Time-Series Foundation Models

Model ReleasesDGX agent

arXiv:2608.04695v1 Announce Type: cross Abstract: Federated adaptation of time-series foundation models (TSFMs) is attractive for building energy forecasting because meter data are private, distribute

Unforgettable Generalization in Language Models

TutorialsDGX agent

arXiv:2409.02228v2 Announce Type: replace-cross Abstract: When language models (LMs) are trained to forget (or 'unlearn'') a skill, how precisely does their behavior change? We study the behavior of t

5 Aug 2026

Adaptive Two-Stage Visual Token Pruning for Efficient Inference in Video-Language Models

Model ReleasesDGX agent

arXiv:2608.03112v1 Announce Type: cross Abstract: Vision-language models excel at image and video understanding but suffer from high inference latency due to the need to process thousands of tokens pe

Anthropic confirms it is building an in-house silicon team to design custom chips for Claude, co-designing hardware and models and using a 'multi-chip approach' (Tom Carter/Business Insider)

Model ReleasesDGX agent

Tom Carter / Business Insider: Anthropic confirms it is building an in-house silicon team to design custom chips for Claude, co-designing hardware and models and using a “multi-chip approach” — - Anth

Assessment of Conditional Diffusion Model for Synthetic Histopathology Image Generation

Model ReleasesDGX agent

arXiv:2608.03990v1 Announce Type: new Abstract: Synthetic histopathology image generation has emerged as an approach that may address data scarcity in computational pathology, yet current evaluation m

BanglaWild: An In-the-Wild Bengali Scene Text Recognition Benchmark for OCR and Vision-Language Models

Model ReleasesDGX agent

arXiv:2608.03884v1 Announce Type: cross Abstract: In-the-wild Bengali scene text recognition is largely unmeasured: existing resources target handwritten documents or constrained sign-board parsing, r

CorePath: A Breast-Specialized Pathology Foundation Model for Core Needle Biopsy Diagnosis and Risk-Controlled Report Generation

Model ReleasesDGX agent

arXiv:2608.03079v1 Announce Type: cross Abstract: Breast core needle biopsy (CNB) is central to breast cancer diagnosis yet remains challenging because limited tissue sampling, lesion heterogeneity, a

Disentangling Language Modeling and Boundaries

SafetyDGX agent

arXiv:2608.03599v1 Announce Type: new Abstract: Byte-level language models are usually argued for on the grounds of robustness, multilingual fairness, and character-level skills. We point to a differe

Federated generative event models for tokenized electronic health records

ResearchDGX agent

arXiv:2608.02939v1 Announce Type: new Abstract: Electronic health record foundation models are limited by institutionally siloed data and substantial performance degradation under cross-site transfer.

FOUND-AF: Benchmarking ECG Foundation Models for Atrial Fibrillation Detection

ResearchDGX agent

arXiv:2608.03597v1 Announce Type: new Abstract: Atrial fibrillation (AF) is the most common sustained cardiac arrhythmia and is associated with increased risks of stroke, heart failure, and mortality.

HUKUKBERT: Domain-Specific Language Model for Turkish Law

Model ReleasesDGX agent

arXiv:2604.04790v2 Announce Type: replace Abstract: Natural language processing (NLP) advances have powered a generation of LegalTech systems, but Turkish law remains under-served by domain-specific d

Interpreting Black-Box Large Language Models with Sentence-Level Energy Landscapes

Local AiDGX agent

arXiv:2608.02879v1 Announce Type: new Abstract: The widespread adoption of proprietary Large Language Models (LLMs) accessed strictly through closed APIs has created a critical challenge for responsib

Modeling Long-Term Memory and Temporal Attention Shifts for Video Salient Object Ranking with a New Benchmark

Model ReleasesDGX agent

arXiv:2203.17257v2 Announce Type: replace Abstract: Salient Object Ranking (SOR) aims to estimate the relative saliency order among multiple salient objects. While SOR has been extensively studied in

UniNav: A Unified World-Action Diffusion Model for Visual Navigation

Model ReleasesDGX agent

arXiv:2608.03244v1 Announce Type: new Abstract: Image-goal visual navigation is a fundamental capability for embodied agents. Existing navigation policies efficiently predict waypoint trajectories but

VIBE: A VAD-Informed Benchmark for Entity-Centered Affective Profiling of Large Language Model Outputs

Model ReleasesDGX agent

arXiv:2608.03810v1 Announce Type: cross Abstract: Large language models routinely describe socially salient targets, including political figures, countries, religions, organizations, historical events

4 Aug 2026

All models are currently 20% discounted in Portal, other than GPT-5.6 Terra and Luna which 50% off and DeepSeek V4 Flash 0731 which is 90% o…

Model ReleasesDGX agent

All models in the Portal are discounted by 20 %, except GPT‑5.6 Terra and Luna (50 % off) and DeepSeek V4 Flash 0731 (90 % off). The latest Alibaba Qwen release, Qwen3.8‑Max, is now available for Herm

Attention-Steered Vision-Language Models for Sign Language Translation

ResearchDGX agent

arXiv:2608.00235v1 Announce Type: new Abstract: Vision-language models (VLMs) have emerged as a powerful framework for multimodal video understanding. However, they remain limited in the sign language

CopyCat: Improving Fine-Grained Subject Consistency in Subject-to-Image Models within Seconds

ResearchDGX agent

arXiv:2608.00674v1 Announce Type: new Abstract: Recent subject-to-image models have achieved impressive progress in personalized image generation, yet they still struggle to preserve fine-grained subj

Disentangling Visuo-Tactile Foresight: Oracle-Guided Interface Discovery for World Action Models

Model ReleasesDGX agent

arXiv:2608.00547v1 Announce Type: new Abstract: Contact-rich manipulation remains challenging because successful control depends on physical interaction cues that are often weakly observable from visi

DynImmune-BERT: Dynamic Immune Repertoire Modeling with Neural ODE Driven Continuous Transformers

Model ReleasesDGX agent

arXiv:2607.17244v2 Announce Type: replace Abstract: Longitudinal T cell receptor repertoires contain signals of clonal expansion, contraction, disappearance, and reappearance after immune perturbation

Evolutionary Curriculum Learning Improves Biological Sequence Modeling

SafetyDGX agent

arXiv:2608.00697v1 Announce Type: cross Abstract: Variational autoencoders (VAEs) trained on multiple sequence alignments (MSAs) have emerged as powerful generative models for biological sequences, wi

Expert-Choice Routing Enables Adaptive Computation in Diffusion Language Models

SafetyDGX agent

arXiv:2604.01622v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) enable parallel, non-autoregressive text generation, yet existing DLM mixture-of-experts (MoE) models inherit

HAFI-VLM: A Frequency Perspective for Diagnosing and Enhancing Visual Perception in Vision-Language Models

ResearchDGX agent

arXiv:2608.02124v1 Announce Type: cross Abstract: Vision-language models (VLMs) remain unreliable when predictions require fine-grained visual evidence. We identify a previously overlooked cause: spec

🛡️Introducing Shieldstral, Mistral’s 3B open-weights model for content safety that can be deployed on-device 🧵 http://mistral.ai/news/shie…

Model ReleasesDGX agent

Mistral AI introduced Shieldstral, a 3‑billion‑parameter, open‑weights model designed for content‑safety tasks and capable of on‑device deployment. The announcement was shared via a tweet from @Mistra

IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models

ResearchDGX agent

arXiv:2607.15732v2 Announce Type: replace Abstract: Visual grounding with multimodal large language models is commonly formulated as autoregressive coordinate generation, where a model outputs boundin

Kilobyte Models: Neural Networks as a Seed and a Quantized Latent

Model ReleasesDGX agent

arXiv:2608.00860v1 Announce Type: new Abstract: The cost of storing and transmitting a trained neural network scales with its parameter count, a bottleneck for over-the-air updates, on-device librarie

← Previous
1…9899100101102…1009
Next →