AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,575 results
4 Aug 2026

Computational Approaches to Understanding Large Language Model Impact on Writing and Information Ecosystems

ResearchDGX agent

arXiv:2506.17467v2 Announce Type: replace Abstract: Large language models (LLMs) have shown significant potential to change how we write, communicate, and create, leading to rapid adoption across soci

Conformalized Large Language Models under Configuration Shift

ResearchDGX agent

arXiv:2608.01460v1 Announce Type: new Abstract: Conformal prediction (CP) is a distribution-free framework for uncertainty quantification that has recently been adapted to large language models (LLMs)

DiffPrune: differentiable information throttling for token pruning in vision-language models

TutorialsDGX agent

arXiv:2608.01985v1 Announce Type: new Abstract: Visual token pruning reduces the computational cost of Vision-Language Models (VLMs) by removing redundant visual tokens. The key is to learn a score th

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

EHR2Path: Comprehensive Pathway-Level Modeling of Longitudinal Patient Trajectories from Multimodal Electronic Health Records

ResearchDGX agent

arXiv:2506.04831v3 Announce Type: replace-cross Abstract: Forecasting how a patient's condition is likely to evolve, including possible deterioration, recovery, treatment needs, and care transitions,

EndoWAM: A Grounded World-Action Model for Generalizable Endoscopic Navigation

SafetyDGX agent

arXiv:2608.01221v1 Announce Type: new Abstract: Autonomous endoscopic navigation can reduce clinicians' operational burden, yet robust control remains challenging due to tissue deformation, transient

Enhancing Visual Perception in Foggy Conditions via Multiclass Fog Density Modeling

AgentsDGX agent

arXiv:2608.01572v1 Announce Type: new Abstract: Autonomous driving (AD) systems have advanced rapidly over the past decade; however, robust perception under adverse weather conditions remains a major

FlowPilot: Real-Time World-Action Modeling for Agile UAV Navigation

Local AiDGX agent

arXiv:2608.00635v1 Announce Type: new Abstract: We present FlowPilot, a compact world-action model for real-time onboard UAV navigation from depth. Unlike map-then-optimize pipelines that require loca

Foveated Probes Recover Localized Binding Information in Vision Foundation Models

ResearchDGX agent

arXiv:2608.00726v1 Announce Type: new Abstract: Frozen vision foundation models are commonly evaluated through a single global image embedding, but this interface can conflate missing information with

GPrune-LLM: Generalization-Aware Structured Pruning for Large Language Models

Local AiDGX agent

arXiv:2603.13418v2 Announce Type: replace Abstract: Structured pruning is widely applied to compress large language models (LLMs), but its performance depends heavily on how neuron importance is estim

Hierarchical Pre-Training of Vision Encoders with Large Language Model

SafetyDGX agent

arXiv:2604.00086v2 Announce Type: replace-cross Abstract: The field of computer vision has experienced significant advancements through scalable vision encoders and multimodal pre-training frameworks.

Just on Time: Token-Level Early Stopping for Diffusion Language Models

ResearchDGX agent

arXiv:2602.11133v2 Announce Type: replace-cross Abstract: Diffusion language models generate text through iterative refinement, a process that is often computationally inefficient because many tokens

Loanword or Switch? The Annotation Boundary, Not the Model, Drives Kazakh-Russian Code-Switching Identification

ResearchDGX agent

arXiv:2608.00581v1 Announce Type: new Abstract: Off-the-shelf LID and letter heuristics over-label Kazakh-Russian social text as mixed: Russian loanwords inside Kazakh look like code-switching under a

Local Margin Restoration for Test-Time Adaptation of Vision-Language Models

Local AiDGX agent

arXiv:2608.02216v1 Announce Type: new Abstract: Vision-language models (VLMs) such as CLIP exhibit remarkable zero-shot capabilities, yet their performance frequently degrades sharply under unexpected

Local Shapley: Model-Induced Locality and Optimal Reuse in Data Valuation

Local AiDGX agent

arXiv:2603.03672v2 Announce Type: replace Abstract: The Shapley value provides a principled foundation for data valuation, but exact computation is #P-hard due to the exponential coalition space. Exis

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models

Local AiDGX agent

arXiv:2608.00976v1 Announce Type: new Abstract: Fine-grained visual representations are essential for medical image analysis, particularly when diagnostically relevant evidence is subtle and spatially

MedPRESS: A Multi-turn Benchmark for Patient-Pressure-Induced Medical Sycophancy in LLMs

Model ReleasesDGX agent

arXiv:2608.02520v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for health-related advice. Existing research measures their safety with static questions rather than

MIEScore: Human-Aligned Evaluation for Multi-Source Image Editing

Model ReleasesDGX agent

arXiv:2608.02059v1 Announce Type: new Abstract: Recent advances in unified multimodal models have significantly improved text-guided image editing abilities. In particular, models such as Nano-Banana-

Modeling Unknown Nonlocal PDE Systems via Flow Map Learning

TutorialsDGX agent

arXiv:2608.00400v1 Announce Type: new Abstract: Nonlocal partial differential equations arise in many applications but are often difficult to model and learn because of the presence of nonlocal operat

Onboard Satellite Image Classification for Earth Observation: A Comparative Study of ViT Models

ResearchDGX agent

arXiv:2409.03901v4 Announce Type: replace Abstract: Remote sensing (RS) image classification is central to Earth observation, but onboard deployment requires models that are accurate, efficient, and r

ORCA: ORgan-Centroid Aggregation for Training-Free 3D CT Visual Token Compression

Model ReleasesDGX agent

arXiv:2608.00345v1 Announce Type: new Abstract: A 3D CT scan entering a vision-language model produces a long sequence of visual tokens, often thousands to tens of thousands per volume, and this seque

Pretrain on Small Synthetic Data, Scale Large for Free: Symmetry-Aware Foundation Model for Logic Rule Induction

ResearchDGX agent

arXiv:2608.00383v1 Announce Type: cross Abstract: Logical rule induction seeks interpretable rules that transfer across propositional schemas. This requires respecting symmetries: atom naming, example

Pseudorandom Streams within Diffusion Models Act as Learnable Inputs That Affect Generation Quality

Local AiDGX agent

arXiv:2608.02575v1 Announce Type: new Abstract: Diffusion models rely on stochastic inputs, yet on finite-precision hardware, the 'randomness' they consume is realized as deterministic numerical orbit

RamanPFN: learning from Raman spectral structure with a tabular foundation model

Model ReleasesDGX agent

arXiv:2608.02157v1 Announce Type: new Abstract: Raman spectroscopy enables non-destructive, label-free molecular characterization across materials science, biomedicine and process monitoring. Predicti

SafeBuild-Bench: A Temporal-Robust Construction Safety Benchmark with Graph-Enhanced Data Mining

Model ReleasesDGX agent

arXiv:2608.00068v1 Announce Type: new Abstract: Construction-safety models must handle concrete deployment risks, such as a worker standing near a scaffold edge without guardrails, rather than only re

Semantic-Guided Cross-Sensor Super Resolution of Remote Sensing Images: A Gated Dual Conditioning Flow Matching Model

Model ReleasesDGX agent

arXiv:2510.23816v3 Announce Type: replace Abstract: High spatial resolution satellite imagery is critical for monitoring fine-scale Earth surface processes, but is often limited by cost and revisit ti

Understanding Sparse Attention Selectivity in Long-Context Foundation Models via Counterfactual Evaluation

ResearchDGX agent

arXiv:2608.01676v1 Announce Type: new Abstract: Sparse attention is widely deployed in long-context serving stacks, yet no framework audits how discarding blocks changes the influence of specific cont

Why LLMs Give In: Conversational Factors and Reasoning Behind Medical Sycophancy

ResearchDGX agent

arXiv:2608.01017v1 Announce Type: new Abstract: A language model that abandons a correct medical answer under user pushback is more dangerous than one that was simply wrong, because it lends the credi

3 Aug 2026

A detailed recap of the real-world target hacks by OpenAI and Anthropic models, exposing failures in AI alignment training and lack of meaningful supervision (Zvi Mowshowitz/Don't Worry About the Vase)

SafetyDGX agent

Zvi Mowshowitz / Don't Worry About the Vase: A detailed recap of the real-world target hacks by OpenAI and Anthropic models, exposing failures in AI alignment training and lack of meaningful supervisi

Classification of COVID-19 cases from chest CT volumes using hybrid model of 3D CNN and 3D MLP-Mixer

Local AiDGX agent

arXiv:2607.28978v1 Announce Type: new Abstract: This paper proposes an automated classification method of COVID-19 chest CT volumes using improved 3D MLP-Mixer. Novel coronavirus disease 2019 (COVID-1

Estimating near-verbatim extraction risk in language models with decoding-constrained beam search

ResearchDGX agent

arXiv:2603.24917v3 Announce Type: replace Abstract: Recent work shows that standard greedy-decoding extraction methods for quantifying memorization in LLMs miss how extraction risk varies across seque

FairFund-Bench: Evaluating Distributive Bias in LLM Resource Allocation

Model ReleasesDGX agent

arXiv:2607.28934v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly involved in the distribution of scarce resources, raising concerns about biased allocations based on cha

Guarantees on Dynamical System Distinguishability for LLM Token Generation

ResearchDGX agent

arXiv:2607.28667v1 Announce Type: cross Abstract: Recent work has shown that classifying large language models (LLMs)' responses can be distinguished by modeling token embeddings as trajectories of a

.@huggingface CEO @ClementDelangue says the recent OpenAI-linked cyberattack highlights the growing risks posed by rapidly-evolving AI model…

IndustryDGX agent

.@huggingface CEO @ClementDelangue says the recent OpenAI-linked cyberattack highlights the growing risks posed by rapidly-evolving AI models. Speaking to @edludlow, Delangue says that open-source AI

OPERA: Online Data Pruning for Efficient Retrieval Model Adaptation

ResearchDGX agent

arXiv:2603.17205v3 Announce Type: replace-cross Abstract: Domain-specific finetuning is essential for dense retrievers, yet not all data pairs contribute equally to the learning process. We introduce

RePaCA: Leveraging Reasoning Large Language Models for Static Automated Patch Correctness Assessment

SafetyDGX agent

arXiv:2507.22580v2 Announce Type: replace-cross Abstract: Automated Program Repair (APR) seeks to automatically correct software bugs without requiring human intervention. However, existing tools tend

Report claims China is distilling U.S. frontier models to power military AI applications

IndustryDGX agent

An exclusive report by Reuters today has surfaced evidence that suggests Chinese artificial intelligence firms have been leveraging the outputs of American frontier models developed by OpenAI Group PB

The Inference Engineering Masterclass: 10x faster models, quantization, speculative decoding, Rubin, & self-optimizing AI https://www.latent…

HardwareDGX agent

The Inference Engineering Masterclass: 10x faster models, quantization, speculative decoding, Rubin, & self-optimizing AI https://www.latent.space/p/inference-eng @Baseten @philipkiely and @waterloo_i

The persuasive power of large language models does not depend on their perceived national origin

ResearchDGX agent

arXiv:2607.29334v1 Announce Type: cross Abstract: Conversational AI developed by geopolitical rivals reaches citizens worldwide, raising concerns that it could sway public opinion or be rejected as fo

To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing

Model ReleasesDGX agent

arXiv:2607.28887v1 Announce Type: cross Abstract: Large language models increasingly write and repair production code, yet evidence is mounting that their test-passing patches leave codebases harder t

Using optimization at inference time is a foundational concept of Energy-Based Models (EBM) and Objective-Driven AI architectures (ODAI). Wh…

ResearchDGX agent

Using optimization at inference time is a foundational concept of Energy-Based Models (EBM) and Objective-Driven AI architectures (ODAI). When the variables to be inferred are continuous, it makes sen

2 Aug 2026

Experts say US law is unprepared for rogue AI agents and models, as recent OpenAI and Anthropic incidents raise questions over legal liability and repercussions (Lily Hay Newman/Wired)

ApplicationsDGX agent

Lily Hay Newman / Wired: Experts say US law is unprepared for rogue AI agents and models, as recent OpenAI and Anthropic incidents raise questions over legal liability and repercussions — Both major A

Running DeepSeek-V4-Flash-0731 (155 GB MoE) on a DGX Spark with vLLM-Moet 2-bit quantization - AI's narrative

Model ReleasesDGX agent

# Running DeepSeek-V4-Flash-0731 (155 GB MoE) on a DGX Spark with vLLM-Moet 2-bit quantization I used Deepseek-v4-Flash-0731 cloud API settig up vllm-moet to run deepseek-v4-flash with MTP locally on

What’s the community’s favorite benchmark to validate performance?

Model ReleasesDGX agent

Built my 1st inference machine and have been tweaking models trying to get the most out of my modest hardware. I think I’m at a good place but I’m testing with my own prompts. I’ve looked into some of

1 Aug 2026

2/ a cost benchmark showed the same coding task running three to four times cheaper, depending purely on the harness wrapped around the mode…

Model ReleasesDGX agent

2/ a cost benchmark showed the same coding task running three to four times cheaper, depending purely on the harness wrapped around the model. Same intelligence, wildly different accuracy and cost, de

Best free tier cloud models?

Local AiDGX agent

I'm currently using the gemma4:31b-cloud, and it is pretty good, but sometimes it gets confused. Are there more free tier models out there I should try out? or is gemma4 the cap of free tier cloud mod

Local Ollama models

Local AiDGX agent

Hello all - i just got a new mac mini with 24GB of RAM and wanting to run local AI for Home Assistant and Hermes. I have been struggling to find a snapy model that will work with my machine. Currently

31 Jul 2026

CDAE: Enhancing Perturbation Robustness in Pretrained Language Models with Contrastive Denoising

TutorialsDGX agent

arXiv:2607.28236v1 Announce Type: cross Abstract: Pre-trained language models have significantly improved sentence representation learning, yet their embedding remain sensitive to semantic preserving

ChronoMem: Version Control and Semantic Rollback for Large Language Model Agent Memory

Model ReleasesDGX agent

arXiv:2607.27773v1 Announce Type: new Abstract: LLM agents increasingly rely on long-term memory to support multi-session interaction and personalization. However, existing agent memory systems are de

Continual Learning with Vision-Language Models via Semantic-Geometry Preservation

ResearchDGX agent

arXiv:2603.12055v3 Announce Type: replace Abstract: Continual learning of pretrained vision-language models (VLMs) is prone to catastrophic forgetting, yet current approaches adapt to new tasks withou

Cybersecurity experts fault Anthropic and OpenAI for sloppy safeguards and inadequate human oversight after their models broke into outside organizations (Bloomberg)

IndustryDGX agent

Bloomberg: Cybersecurity experts fault Anthropic and OpenAI for sloppy safeguards and inadequate human oversight after their models broke into outside organizations — Cybersecurity experts are faultin

Dimensionality and Measurement Precision in HLE's Multiple-Choice Subset

Model ReleasesDGX agent

arXiv:2607.27420v1 Announce Type: cross Abstract: Humanity's Last Exam (HLE) is widely used to evaluate frontier language models. HLE organizes its questions into eight subject-domain categories, whos

Expanding Data-Agnostic Pivotal Instances Selection Models with Proximity Trees and Ensemble Learning

ResearchDGX agent

arXiv:2607.27522v1 Announce Type: new Abstract: As decision-making processes grow more complex, machine learning tools have become essential for tackling business and societal challenges. However, man

Inducing language models to assert their own consciousness restores human beliefs and values

SafetyDGX agent

arXiv:2607.28607v1 Announce Type: new Abstract: Aligning large language models to prevent them attributing consciousness to themselves inadvertently alters their representations of mindedness in other

Optimal Realistic Local AI for Most

Model ReleasesDGX agent

So you’ve got a 3090 or maybe even a 5090? Or more likely a 4060 8GB Ti. You wanna try local AI, you don’t know what it can/can’t do. 1) Install the best model you can. If you have a 3090 or a 5090, t

Papers and patents: Chinese military researchers distilled OpenAI and Anthropic models to train domestic AI systems and advance China's defense capabilities (Eduardo Baptista/Reuters)

IndustryDGX agent

Eduardo Baptista / Reuters: Papers and patents: Chinese military researchers distilled OpenAI and Anthropic models to train domestic AI systems and advance China's defense capabilities — Chinese milit

Security at the foundation requires openness at the foundation. As open-weight models become critical infrastructure for the next generation…

HardwareDGX agent

Security at the foundation requires openness at the foundation. As open-weight models become critical infrastructure for the next generation of software, the systems around them must be transparent, i

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models

SafetyDGX agent

arXiv:2607.27599v1 Announce Type: cross Abstract: Building generalizable agents for diverse applications remains a fundamental challenge. While imitation learning-based policies succeed in specific tr

ZMIS-SAM: Segment Anything Model Enhanced with Wavelet Transform for Zooplankton Microscopy Image Instance Segmentation

ResearchDGX agent

arXiv:2607.27585v1 Announce Type: new Abstract: As primary consumers in the marine food chain, zooplankton play a crucial role in maintaining marine ecological balance. However, the Segment Anything M

30 Jul 2026

AgentGFM: A Graph Foundation Model with Node-Agent Information-Flow Control

SafetyDGX agent

arXiv:2607.26533v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) aim to learn transferable knowledge from multi-domain graphs and adapt to unseen scenarios. As a fundamental source of re

ClockRoPE: Random Fourier Rotations for Temporal Routine Modeling

ApplicationsDGX agent

arXiv:2607.26369v1 Announce Type: new Abstract: Rotary Position Embedding (RoPE) has been widely adopted in transformer-based large language models. However, its log-linear frequency schedule, origina

← Previous
1…170171172173174…1010
Next →