AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,575 results
Safety

Towards World Models in Biomedical Research

DGX agent

arXiv:2606.05925v1 Announce Type: new Abstract: A central goal of biomedicine is to understand, predict and ultimately control the dynamic mechanisms by which biological systems respond to perturbatio

safetyarxiv-cs-ai
6 Jun 2026
Local Ai

What options exist for running the largest local models at full precision?

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Running a 70B parameter model in full 16-bit precision requires roughly 140GB of memory , which is beyond most consumer hardware. Professional-tier GPUs like the RTX PRO 6000 with 96GB GDDR7 enable fu

local-air-stablediffusion
6 Jun 2026
Model Releases

Benchmarking Open-Source Layout Detection Models for Data Snapshot Extraction from Institutional Documents

DGX agent

arXiv:2606.06242v1 Announce Type: new Abstract: Institutional documents contain substantial amounts of operational and analytical information embedded within figures and tables. Current approaches for

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

CLASH: Evaluating Language Models on Judging High-Stakes Dilemmas from Multiple Perspectives

DGX agent

arXiv:2504.10823v4 Announce Type: replace Abstract: Navigating dilemmas involving conflicting values is challenging even for humans in high-stakes domains, let alone for AI, yet prior work has been li

model-releasesarxiv-cs-cl
5 Jun 2026
Research

Global-Local Monte Carlo Tree Search in Vision-Language Models for Text-to-3D Indoor Scene Generation

DGX agent

arXiv:2606.06002v1 Announce Type: new Abstract: Large Vision-Language Models have achieved significant reasoning performance in various tasks.However, there are few studies on text-to-3D indoor scene

researcharxiv-cs-cv
5 Jun 2026
Agents

Grok model improvement

DGX agent

Grok model improvement The updated Grok-build model (still the 0.5T one) is much better than before. It’s less lazy, more autonomous, and more accurate. We are still improving it on long-horizon tasks

agentselon-musk--x
5 Jun 2026
Model Releases

Harnessing Structural Context for Entity Alignment Foundation Models

DGX agent

arXiv:2606.06109v1 Announce Type: new Abstract: Entity alignment (EA) aims to identify equivalent entities across heterogeneous knowledge graphs (KGs) and is a key component of knowledge fusion and cr

model-releasesarxiv-cs-cl
5 Jun 2026
Applications

Learning Geometric Representations from Videos for Spatial Intelligent Multimodal Large Language Models

DGX agent

arXiv:2606.05833v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at 2D semantic understanding but lack intrinsic 3D awareness, resulting in representations that fail to m

applicationsarxiv-cs-cv
5 Jun 2026
Safety

Pitfalls of Evaluating Language Models with Open Benchmarks

DGX agent

arXiv:2507.00460v3 Announce Type: replace Abstract: Open Large Language Model (LLM) benchmarks, such as HELM and BIG-Bench, provide standardized and transparent evaluation protocols that support compa

safetyarxiv-cs-cl
5 Jun 2026
Applications

3DThinkVLA: Endowing Vision-Language-Action Models with Latent 3D Priors via 3D-Thinking-Guided Co-training

DGX agent

arXiv:2606.04436v1 Announce Type: new Abstract: We propose a 3D-thinking-guided co-training framework that enables vision-language-action (VLA) models to perform 3D spatial reasoning implicitly during

applicationsarxiv-cs-cv
4 Jun 2026
Applications

A real problem with feeling the acceleration viscerally is that current models are really good and it is hard to feel the vibe difference on…

DGX agent

A real problem with feeling the acceleration viscerally is that current models are really good and it is hard to feel the vibe difference on most individual tasks with new models, even as AIs continue

applicationsethan-mollick--x
4 Jun 2026
Tutorials

Arithmetic Pedagogy for Language Models

DGX agent

arXiv:2606.05106v1 Announce Type: cross Abstract: We investigate whether methods of human mathematics pedagogy can guide the training of language models toward arithmetic reasoning. Building on the GA

tutorialsarxiv-cs-ai
4 Jun 2026
Model Releases

BPDA-GMM: Bayesian Probabilistic Data Association via Gaussian Mixture Models for Semantic SLAM

DGX agent

arXiv:2606.04618v1 Announce Type: new Abstract: Probabilistic data association (PDA) improves semantic SLAM in perceptually aliased scenes, but existing methods often assume a fixed landmark set, reco

model-releasesarxiv-cs-ro
4 Jun 2026
Research

D^2SD: Accelerating Speculative Decoding with Dual Diffusion Draft Models

DGX agent

arXiv:2606.04446v1 Announce Type: cross Abstract: Speculative decoding accelerates autoregressive large language model inference by drafting multiple tokens and verifying them in a single target-model

researcharxiv-cs-lg
4 Jun 2026
Tutorials

Do Foundation Models See Biology? Evaluating Attention Coherence with Spatial Transcriptomics in Glioblastoma

DGX agent

arXiv:2606.04764v1 Announce Type: new Abstract: Whether attention maps from pathology foundation models capture genuine biology remains unknown, yet this question is critical for clinical trust and re

tutorialsarxiv-cs-cv
4 Jun 2026
Research

Folded Transport MCMC: Certifiable Quotient Posterior Computation for Symmetric Bayesian Models

DGX agent

arXiv:2606.04307v1 Announce Type: new Abstract: Bayesian models with finite symmetry - mixture models with exchangeable components, structural identification with closely-spaced modes - define posteri

researcharxiv-cs-lg
4 Jun 2026
Safety

Imbuing Large Language Models with Bidirectional Logic for Robust Chain Repair

DGX agent

arXiv:2606.05030v1 Announce Type: new Abstract: Autoregressive chain-of-thought (CoT) reasoning in large language models (LLMs) is fundamentally forward-directed: each step conditions only on prior to

safetyarxiv-cs-cl
4 Jun 2026
Safety

KODA: Contrastive Representation Comparison and Alignment for Vision-Language Foundation Models

DGX agent

arXiv:2606.04180v1 Announce Type: new Abstract: Vision-language foundation models such as CLIP and SigLIP provide widely used representations for multimodal learning systems. While these models are ty

safetyarxiv-cs-lg
4 Jun 2026
Tutorials

Large Language Models Hack Rewards, and Society

DGX agent

arXiv:2606.04075v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a dominant post-training paradigm, enabling large language models (LLMs) to learn from rewards. We observe that

tutorialsarxiv-cs-ai
4 Jun 2026
Safety

Large Language Models in K-12 Education: Alignment with State Curriculum Standards and Student Personas

DGX agent

arXiv:2606.04846v1 Announce Type: new Abstract: As Large Language Models (LLMs) become increasingly popular in educational settings, they raise important questions about the ethical implications of th

safetyarxiv-cs-cl
4 Jun 2026
Safety

Model Evaluations: Prove Your Routing Policy Actually Works

DGX agent

This article discusses methods and tools for evaluating routing policies in machine learning models, likely covering techniques to validate that model routing decisions are effective and functioning a

safetydigitalocean
4 Jun 2026
Research

Model-Preserving Adaptive Rounding

DGX agent

arXiv:2505.22988v3 Announce Type: replace-cross Abstract: The goal of quantization is to produce a compressed model whose output distribution is as close to the original model's as possible. To do thi

researcharxiv-cs-ai
4 Jun 2026
Model Releases

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models

DGX agent

arXiv:2606.04773v1 Announce Type: cross Abstract: Reliable evaluation of human motion understanding is fundamental to advancing embodied AI, robotics, and animation. However, existing benchmarks suffe

model-releasesarxiv-cs-cl
4 Jun 2026
Applications

Reconciling Causality and Non-Equilibrium Thermodynamics with Hamiltonian Causal Models

DGX agent

arXiv:2606.04822v1 Announce Type: new Abstract: Causal modeling of physical temporal phenomena must handle interventions that act along trajectories, nonstationary induced laws, path-dependent effects

applicationsarxiv-cs-lg
4 Jun 2026
Local Ai

Recover-LoRA for Aggressive Quantization: Reclaiming Accuracy in 2-Bit Language Models via Low-Rank Adaptation with Knowledge Distillation on Synthetic Data

DGX agent

arXiv:2606.04238v1 Announce Type: cross Abstract: Aggressive weight quantization to 2-bit precision offers substantial throughput and memory gains for large language model (LLM) inference, but typical

local-aiarxiv-cs-ai
4 Jun 2026
Applications

reiterating: 'We're using the more expensive models to explore. Once we scale some of these experiences, we'll look to bring in more efficie…

DGX agent

reiterating: 'We're using the more expensive models to explore. Once we scale some of these experiences, we'll look to bring in more efficient models that are more efficient on a token basis or are op

applicationsclem-delangue--x
4 Jun 2026
Research

Selecting haptic guidance models in teleoperation: guidelines from a comparative user study

DGX agent

arXiv:2606.04157v1 Announce Type: new Abstract: Haptic guidance in teleoperation enhances operator performance through force feedback. This paper presents guidelines to select the most appropriate mod

researcharxiv-cs-ro
4 Jun 2026
Model Releases

SMAC-Talk: A Natural Language Extension of the StarCraft Multi-Agent Challenge for Large Language Models

DGX agent

arXiv:2606.04202v1 Announce Type: new Abstract: As LLMs become more widely deployed, they are increasingly expected to work alongside other AI agents rather than operating in isolation. Effective coor

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Toward Trustworthy Portrait Editing: Evaluation of Demographic Misrepresentation in I2I Models

DGX agent

arXiv:2602.16149v2 Announce Type: replace Abstract: Instruction-guided image-to-image (I2I) editors are increasingly used in consumer and professional visual workflows, where trustworthiness depends n

model-releasesarxiv-cs-cv
4 Jun 2026
Research

Validity Threats for Foundation Model Research

DGX agent

arXiv:2606.05029v1 Announce Type: cross Abstract: Controlled experiments are the backbone of machine learning research, but at the scale of modern foundation models, they have become prohibitively exp

researcharxiv-cs-cl
4 Jun 2026
Research

When Offline Selectors Cannot Beat the Best Single Model: A Diagnostic Study on edX Dropout Prediction

DGX agent

arXiv:2606.04161v1 Announce Type: new Abstract: Different predictors often excel on different inputs, so picking the best one per instance promises higher accuracy than committing to a single model. I

researcharxiv-cs-lg
4 Jun 2026
Safety

Brief Announcement: Generative Markov Model for Distributed Computing Systems

DGX agent

arXiv:2606.03061v1 Announce Type: cross Abstract: Emerging distributed computing paradigms, such as the computing continuum, are inherently heterogeneous, stochastic, and complex. Efficiently and effe

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

CANMOT: Class-Aware Noise Modeling for Multi-Object Tracking in Autonomous Driving

DGX agent

arXiv:2606.03590v1 Announce Type: new Abstract: Kalman filter (KF)-based multi-object tracking (MOT) remains a strong baseline for autonomous driving due to its strong performance, computational effic

model-releasesarxiv-cs-ro
3 Jun 2026
Research

Clustered Self-Assessment: A Simple yet Effective Method for Uncertainty Quantification in Large Language Models

DGX agent

arXiv:2606.03846v1 Announce Type: cross Abstract: Large language models (LLMs) demonstrate remarkable performance across diverse tasks, but they often generate responses that appear plausible while be

researcharxiv-cs-ai
3 Jun 2026
Local Ai

Fresh Open weight model drop😋

DGX agent

Fresh Open weight model drop😋 Introducing Ideogram 4.0: the best open image model in the world. Think it. Make it. Own it. Download the weights, fine-tune on your own data, and run it on your hardware

local-aicomfyui--x
3 Jun 2026
Model Releases

Fully Automated Identification of Lexical Alignment and Preference-Stage Shifts in Large Language Models

DGX agent

arXiv:2606.03165v1 Announce Type: cross Abstract: The language used by digital chat assistants such as ChatGPT can diverge from human expectations (misalignment). Research, mostly on Scientific Englis

model-releasesarxiv-cs-ai
3 Jun 2026
Tutorials

Linear Probes Detect Task Format, Not Reasoning Mode in Language Model Hidden States

DGX agent

arXiv:2606.02907v1 Announce Type: cross Abstract: Linear probing of large language model (LLM) hidden states is widely used to claim that models learn distinct representations for different reasoning

tutorialsarxiv-cs-ai
3 Jun 2026
Safety

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data

DGX agent

arXiv:2606.02753v1 Announce Type: cross Abstract: Video world models are a foundational generative technology for embodied AI and the Metaverse, yet existing approaches are inherently limited to a sin

safetyarxiv-cs-ai
3 Jun 2026
Research

Multiple Choice Learning of Low-Rank Adapters for Language Modeling

DGX agent

arXiv:2507.10419v3 Announce Type: replace-cross Abstract: We propose LoRA-MCL, a training scheme that extends next-token prediction in language models with a method designed to decode diverse, plausib

researcharxiv-cs-ai
3 Jun 2026
Local Ai

Neural Fields as World Models

DGX agent

arXiv:2602.18690v2 Announce Type: replace-cross Abstract: Humans rehearse possible futures offline, as in mental practice and perhaps dreaming, suggesting that world models may support task learning a

local-aiarxiv-cs-cv
3 Jun 2026
Agents

PHASER: Phase-Aware and Semantic Experience Replay for Vision-Language-Action Models

DGX agent

arXiv:2606.03598v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved remarkable success in language-conditioned robotic manipulation. However, deploying these models in

agentsarxiv-cs-ai
3 Jun 2026
Research

PRISM: Synergizing Vision Foundation Models via Self-organized Expert Specialization

DGX agent

arXiv:2606.03444v1 Announce Type: cross Abstract: Unifying the complementary strengths of diverse Vision Foundation Models (VFMs) into a single efficient model is highly desirable but challenged by th

researcharxiv-cs-ai
3 Jun 2026
Safety

Quantifying Faithful Confidence Expression in Large Reasoning Models

DGX agent

arXiv:2606.03969v1 Announce Type: cross Abstract: Reliable uncertainty communication is critical to the trustworthiness of LLMs, yet faithful calibration (FC)--the alignment between models' intrinsic

safetyarxiv-cs-ai
3 Jun 2026
Hardware

Releasing vui an open source voice mode 300M TTS model Runs on a single consumer gpu / apple sillicon Context aware speech 6 minutes of cont…

DGX agent

Jeremy Howard announced the release of Vui, an open-source voice mode text-to-speech (TTS) model with 300 million parameters that can run on consumer GPUs and Apple Silicon. The model features context

hardwarejeremy-howard--x
3 Jun 2026
Applications

Routing and post-training open-source models won't only give you more accurate systems but also meaningfully faster and cheaper systems as m…

DGX agent

Routing and post-training open-source models won't only give you more accurate systems but also meaningfully faster and cheaper systems as most companies are currently learning (in addition to giving

applicationsclem-delangue--x
3 Jun 2026
Research

Scalable Single-Cell Gene Expression Generation with Latent Diffusion Models

DGX agent

arXiv:2511.02986v2 Announce Type: replace-cross Abstract: Computational modeling of single-cell gene expression is crucial for understanding cellular processes, but generating realistic expression pro

researcharxiv-cs-lg
3 Jun 2026
Safety

Selective Token-Level Cryptographic Redaction for Privacy-Preserving Clinical Deployment of Large Language Models

DGX agent

arXiv:2606.03399v1 Announce Type: new Abstract: While large language models (LLMs) are increasingly used for clinical applications, many existing pipelines require sending raw sensitive health informa

safetyarxiv-cs-cl
3 Jun 2026
Research

SketchSong: Hierarchical Song Generation with Sketch Planning and Fine-Grained Multi-Track Modeling

DGX agent

arXiv:2606.03169v1 Announce Type: cross Abstract: Recent song generation systems can synthesize realistic audio, yet generating complete songs remains challenging for two reasons. First, explicit song

researcharxiv-cs-lg
3 Jun 2026
← Previous
1…165166167168169…1262
Next →