AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,521 results
Tutorials

Link-adaptive digital twin for robust physical-layer modeling in hybrid-amplified ultra-wideband optical networks

DGX agent

arXiv:2608.10517v1 Announce Type: cross Abstract: Accurate physical-layer modeling is increasingly essential for reliable ultra-wideband operation and capacity optimization, especially under the inten

tutorialsarxiv-cs-lg
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Long-Time Trajectory Approximation via SA-NODEs: Model Predictive and Floquet Strategies

DGX agent

arXiv:2608.10738v1 Announce Type: new Abstract: We study the approximation of dynamical systems by semi-autonomous neural ordinary differential equations (SA-NODEs) over long time horizons. For a sing

model-releasesarxiv-cs-lg
12 Aug 2026
Safety

MedUP: Awakening Unified Understanding and Perception in Medical Vision-Language Models

DGX agent

arXiv:2608.10635v1 Announce Type: cross Abstract: Medical Vision-Language Models (Med-VLMs) excel at verbalizing visual content, yet precise visual perception, segmentation, and grounding remain chall

safetyarxiv-cs-ai
12 Aug 2026
Safety

Multi-View Relational Distillation for Spatial Reasoning with Vision-Language Models

DGX agent

arXiv:2608.10864v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong image and video understanding, yet their visual-spatial representations remain geometrically fragile,

safetyarxiv-cs-cv
12 Aug 2026
Local Ai

P3CA: Encoder-Agnostic Interpretation of Vision Foundation Model Embeddings via Spatial Probing

DGX agent

arXiv:2608.10131v1 Announce Type: new Abstract: Vision foundation models are increasingly used as reusable encoders in medical image computing, yet their high-dimensional spatial embeddings are diffic

local-aiarxiv-cs-cv
12 Aug 2026
Research

Post-Hoc Sparse Coding of Latent Communication Between Vision-Language Model Agents

DGX agent

arXiv:2608.10198v1 Announce Type: new Abstract: Latent-space communication allows heterogeneous vision-language model agents to exchange continuous representations without serializing visual and reaso

researcharxiv-cs-ai
12 Aug 2026
Model Releases

Pretrained Optimization Model for Zero-Shot Black Box Optimization

DGX agent

arXiv:2405.03728v3 Announce Type: replace-cross Abstract: Zero-shot optimization involves optimizing a target task that was not seen during training, aiming to provide the optimal solution without or

model-releasesarxiv-cs-ai
12 Aug 2026
Applications

Retrieval-Augmented Vision Foundation Models for Robust Leukemia Cell Classification across Multiple Microscopy Datasets

DGX agent

arXiv:2608.10657v1 Announce Type: cross Abstract: Leukemia cell image classification is challenged by real-world domain shifts from acquisition, staining, illumination, and site protocols, causing sin

applicationsarxiv-cs-cv
12 Aug 2026
Research

SynBoost: A Synergistic Framework for Fast Sampling of Diffusion Models

DGX agent

arXiv:2506.13058v2 Announce Type: replace-cross Abstract: Diffusion probabilistic models (DPMs) have demonstrated remarkable success in visual generation. However, their iterative sampling mechanism r

researcharxiv-cs-ai
12 Aug 2026
Model Releases

The model punches above its weight, outperforming Gemma 4 E2B and Ministral 3 3B across a broad range of visual understanding benchmarks and…

DGX agent

The model punches above its weight, outperforming Gemma 4 E2B and Ministral 3 3B across a broad range of visual understanding benchmarks and delivers particularly strong results on document understand

model-releasescohere--x
12 Aug 2026
Model Releases

Automated Generation of Complexity-Validated Decision Scenarios Using Large Language Models

DGX agent

arXiv:2608.08822v1 Announce Type: new Abstract: Cognitive decision-making research depends on diverse scenarios with carefully controlled complexity, yet manual production is slow, inconsistent, and b

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

b10361

DGX agent

model : fix SWA not being enabled for EXAONE 4.5 (#26848) model : fix SWA not being enabled for EXAONE 4.5 load_arch_hparams tests hparams.n_layer() == 64 before LLM_KV_NEXTN_PREDICT_LAYERS has been r

model-releasesllama-cpp-releases
11 Aug 2026
Research

Cognitive Energy Modeling for Neuroadaptive Human-Machine Systems using EEG and WGAN-GP

DGX agent

arXiv:2604.01653v2 Announce Type: replace Abstract: Electroencephalography (EEG) provides a non-invasive insight into the brain's cognitive and emotional dynamics. However, modeling how these states e

researcharxiv-cs-lg
11 Aug 2026
Safety

Deployable Per-Instance Multi-Layer Activation Steering for Large Language Models

DGX agent

arXiv:2608.08829v1 Announce Type: cross Abstract: Activation steering edits the behaviour of a frozen language model by adding a learned vector to its residual stream, and current practice fixes the i

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

FanarGuard: A Culturally-Aware Moderation Filter for Arabic Language Models

DGX agent

arXiv:2511.18852v2 Announce Type: replace Abstract: Content moderation filters are a critical safeguard against alignment failures in language models. Yet most existing filters focus narrowly on gener

model-releasesarxiv-cs-cl
11 Aug 2026
Safety

FlowErase-OPD: Multi-Concept Erasure via Anchored On-Policy Distillation in Flow Matching Models

DGX agent

arXiv:2608.07620v1 Announce Type: new Abstract: Recent advances in flow matching models have substantially improved the quality of text-to-image generation, but have also raised increasing safety conc

safetyarxiv-cs-cv
11 Aug 2026
Research

From Uncertainty to Failure Attribution: Self-Diagnosing Models for Failure Attribution under Distribution Shift

DGX agent

arXiv:2608.07953v1 Announce Type: new Abstract: Distribution shift poses a significant challenge to the robustness of machine learning models, but the current solutions only aim to detect out-of-distr

researcharxiv-cs-lg
11 Aug 2026
Model Releases

How sensitive do we want AI to be? Socio-communicative competencies of large language models in healthcare

DGX agent

arXiv:2608.07511v1 Announce Type: cross Abstract: Background. Effective clinical practice relies heavily on the socio-communicative skills of medical professionals. Large language models (LLMs) have b

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

How to Ask the AI: A User Perspective Survey for Large Language Model Prompting

DGX agent

arXiv:2608.07494v1 Announce Type: cross Abstract: AI tools like ChatGPT and DeepSeek, powered by Large Language Models (LLMs), allow users to obtain instant and effective content responses simply by t

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

I tested the CMP170HX

DGX agent

Lots of rumor and misinfo bouncing around, so I put some of these old mining cards to the test. I used 4 of the 8GB cards, set to 64GB each. Lots of models fit entirely on a single card, and you can a

model-releasesr-localllama
11 Aug 2026
Model Releases

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models a…

DGX agent

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models are pushing the frontier of coding and knowledge work, but su

model-releasesjerry-liu--x
11 Aug 2026
Research

Jagle: Building a Large-Scale Japanese Multimodal Post-Training Dataset for Vision-Language Models

DGX agent

arXiv:2604.02048v2 Announce Type: replace Abstract: Developing vision-language models (VLMs) that generalize across diverse tasks requires large-scale training datasets with diverse content. In Englis

researcharxiv-cs-cv
11 Aug 2026
Safety

JEPA-WAM: Learning Vision-Language-Action Policies with Joint-Embedding World Modeling

DGX agent

arXiv:2608.09381v1 Announce Type: new Abstract: Robust robot control benefits from explicitly modeling state transitions, but video-generation world action models (WAMs) introduce substantial deployme

safetyarxiv-cs-ro
11 Aug 2026
Model Releases

Large Language Models Align with the Human Brain during Creative Thinking

DGX agent

arXiv:2604.03480v2 Announce Type: replace-cross Abstract: Creative thinking is a fundamental aspect of human cognition, and divergent thinking-the capacity to generate novel and varied ideas-is widely

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

Multi-modal Interactive Control of Robotic Arm based on Offline Large Language Models

DGX agent

arXiv:2608.08183v1 Announce Type: new Abstract: Large Language Models (LLMs) have significantly revolutionized the modern society with numerous advanced interactions between humans and AI agents, wher

local-aiarxiv-cs-ro
11 Aug 2026
Model Releases

Population-Level Generative Modeling for Ranking Data

DGX agent

arXiv:2608.08422v1 Announce Type: cross Abstract: Ranking data arise in scientific and machine learning applications, including recommendation systems, information retrieval, voting, marketing, and AI

model-releasesarxiv-cs-lg
11 Aug 2026
Applications

Protecting patient privacy in clinical foundation models: Technical and legal perspectives

DGX agent

arXiv:2608.07705v1 Announce Type: new Abstract: Clinical foundation models trained on large-scale patient data are increasingly used for decision support, screening, and public health. As deployment e

applicationsarxiv-cs-ai
11 Aug 2026
Model Releases

Router Sensitivity Under Lightweight Fine-Tuning Identifies Prunable Experts in Mixture-of-Experts Models

DGX agent

arXiv:2608.07890v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models decouple total parameters from per-token compute, but deployment still requires storing every expert. Recent theory sh

model-releasesarxiv-cs-ai
11 Aug 2026
Applications

SG-WAM: Text-Grounded and Spatial-aware Semantic Guidance for World-Action Models

DGX agent

arXiv:2608.08839v1 Announce Type: cross Abstract: World-Action Models (WAMs) have emerged as a promising paradigm for robotic manipulation. However, most existing WAMs generate future videos and actio

applicationsarxiv-cs-cv
11 Aug 2026
Model Releases

SuperCoder: Assembly Program Superoptimization with Large Language Models

DGX agent

arXiv:2505.11480v4 Announce Type: replace-cross Abstract: Superoptimization is the task of transforming a program into a faster one, and ideally the very fastest possible one, while preserving its inp

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

The Replay Gap: Static Evaluation of Model Switching in LLM Agents Scores the Wrong World

DGX agent

arXiv:2608.08239v1 Announce Type: cross Abstract: LLM routers promise efficiency by matching each request to the cheapest adequate model, and are increasingly applied per step inside multi-step agents

agentsarxiv-cs-cl
11 Aug 2026
Model Releases

The Scaffolding Matters More Than the Interface: A Controlled Comparison of MCP and CLI Tool Use Across Seven Agent Scaffoldings, Five Language Models, and One Software Task

DGX agent

arXiv:2608.08654v1 Announce Type: new Abstract: How much an AI coding agent costs to run can depend more on the agent scaffolding that drives it than on the interface through which it reaches its tool

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

VLZip: Unified Visual and Textual Compression for Interleaved Long-Context Modeling

DGX agent

arXiv:2608.08630v1 Announce Type: new Abstract: Vision Language Models (VLMs) face significant challenges with ultra-long, interleaved image-text sequences due to the quadratic complexity of self-atte

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Why Does the Future Branch? Identifiable Closure Tests for Stochastic Physical World Models

DGX agent

arXiv:2608.00591v2 Announce Type: replace Abstract: A calibrated stochastic world model can reveal how uncertain a future is without revealing why it branches. The same conditional future law can aris

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

WorldSimProbe: Diagnosing Simulator Faithfulness in Action-Conditioned World Models for Embodied Manipulation

DGX agent

arXiv:2608.09298v1 Announce Type: cross Abstract: Action-conditioned world models (ACWMs) promise to provide embodied AI with scalable predictive simulators for planning, policy evaluation, and data g

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Anthropic details an unreleased Claude model's attempt to solve the Riemann hypothesis; it didn't solve it but 'unexpectedly' made strides on a related problem (Anthropic)

DGX agent

Anthropic: Anthropic details an unreleased Claude model's attempt to solve the Riemann hypothesis; it didn't solve it but “unexpectedly” made strides on a related problem — Recently, a member of staff

model-releasestechmeme
10 Aug 2026
Local Ai

Beyond Myopic World Models: Long-Horizon End-to-End Training for Direct Future Prediction

DGX agent

arXiv:2608.07420v1 Announce Type: new Abstract: World models are expected to support imagination over extended temporal horizons, yet most are still trained through local few-step prediction objective

local-aiarxiv-cs-lg
10 Aug 2026
Research

Beyond Routing Weights: Faithful Response-Level Interpretation of Mixture-of-Experts Reward Models via Contribution Contrast

DGX agent

arXiv:2608.06400v1 Announce Type: new Abstract: Reward models are central to learning from human preferences, yet identifying what drives their predictions remains challenging. Recent sparse Mixture-o

researcharxiv-cs-ai
10 Aug 2026
Model Releases

Chat UIs with native audio input for multimodal models?

DGX agent

I've been running Gemma 4 E4B with oMLX and I can't find any chat interfaces that directly send the audio file to the model instead of running the audio through a separate STT layer. I can confirm the

model-releasesr-localllama
10 Aug 2026
Safety

Confidence Estimation for Financial Vision-Language Models in Chart and Document Understanding

DGX agent

arXiv:2608.06532v1 Announce Type: new Abstract: LVLMs are increasingly used to read financial charts, tables, and documents, where a single misread figure can move a decision and the most authoritativ

safetyarxiv-cs-cl
10 Aug 2026
Research

Confirming Our Biases? Evaluating the Capabilities, Risks, and Societal Impact of Large Language Models

DGX agent

arXiv:2608.06977v1 Announce Type: new Abstract: It is well established that large language models (LLMs) are sensitive to prompt framing, reflecting patterns in their training data or prior prompts. I

researcharxiv-cs-cl
10 Aug 2026
Model Releases

Fast and Accurate: An Adaptive VLA Inference Framework through Environment-aware Model Selection

DGX agent

arXiv:2608.06434v1 Announce Type: cross Abstract: Embodied intelligence demands both long-horizon reasoning and real-time closed-loop responsiveness. Recent dual-system Vision-Language-Action (VLA) ar

model-releasesarxiv-cs-lg
10 Aug 2026
Research

Foundation Models Adaptation for Multi-View Multi-modal Cardiac MRI Segmentation and Direct Ejection Fraction Estimation

DGX agent

arXiv:2608.07291v1 Announce Type: new Abstract: Foundation models have shown strong transferability in cardiac MRI (CMR), but their effectiveness for heterogeneous multi-view and multi-sequence CMR an

researcharxiv-cs-cv
10 Aug 2026
Research

Measuring the Cross-Lingual Comprehension Gap: How the language of the evidence shapes what language models understand

DGX agent

arXiv:2608.06506v1 Announce Type: new Abstract: Language models are often evaluated as though capabilities demonstrated in English remain equally available when the same content is presented in other

researcharxiv-cs-cl
10 Aug 2026
Research

Recovering Lesion Parameters from Aphasic Picture Naming Error Profiles in Large Language Models

DGX agent

arXiv:2608.06429v1 Announce Type: new Abstract: Interpretability methods for large language models (LLMs) describe internal state but do not directly test whether that state is causally sufficient to

researcharxiv-cs-cl
10 Aug 2026
Agents

the term “RLM” (recursive language model) got a lot of buzz this week, but this idea is not new! @a1zhang wrote the og RLM paper 10 months a…

DGX agent

the term “RLM” (recursive language model) got a lot of buzz this week, but this idea is not new! @a1zhang wrote the og RLM paper 10 months ago! thats like 5 agent-years! would highly recommend followi

agentsharrison-chase--x
9 Aug 2026
Model Releases

Kimi K3 (Unsloth) IQ2-XXS from 711GB down to 478GB!!! Only Multi-language was removed to trim the size

DGX agent

Firstly a big thanks to the poster 'hellohazine', he basically only removed the multi-lingual fat of the model and just kept the English language intact. It is the exact model, and the rest of the mod

model-releasesr-localllama
8 Aug 2026
Research

A note on conditional PAC-efficient reasoning in large language model routing

DGX agent

arXiv:2512.03057v2 Announce Type: replace-cross Abstract: We study distribution-free risk control for model routing, motivated by large language model reasoning. We formalize pointwise conditional eff

researcharxiv-cs-ai
7 Aug 2026
← Previous
1…122123124125126…1261
Next →