AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,015 results
Research

Edge-aware Decoding for Neural Asymmetric Routing

DGX agent

arXiv:2606.02136v1 Announce Type: new Abstract: Neural asymmetric routing models increasingly encode directionality through matrix representations and asymmetry-aware attention. The final routing acti

researcharxiv-cs-lg
2 Jun 2026
Research

Edge Prediction for Roof Wireframe Reconstruction with Transformers

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.02406v1 Announce Type: new Abstract: This paper presents a competitive solution to the S23DR Challenge 2026, which aims to reconstruct 3D house roof wireframe models from sparse SfM point c

researcharxiv-cs-cv
2 Jun 2026
Research

Effects of Varying LLM Access on Essay Writing Behavior

DGX agent

arXiv:2606.00250v1 Announce Type: cross Abstract: Investigating the degree to which large language models (LLMs) affect teaching and learning in universities can help identify strategies for integrati

researcharxiv-cs-ai
2 Jun 2026
Research

Efficient Synthetic Network Generation via Latent Embedding Reconstruction

DGX agent

arXiv:2606.00934v1 Announce Type: cross Abstract: Network data are ubiquitous across the social sciences, biology, and information systems. Generating realistic synthetic network data has broad applic

researcharxiv-cs-lg
2 Jun 2026
Research

ETC: Extreme Token Compression via Task-aware Visual Information Distillation in VLMs

DGX agent

arXiv:2606.00543v1 Announce Type: new Abstract: In Vision-Language Models (VLMs), high-resolution images produce a large number of visual tokens, resulting in high computational costs and KV-cache ove

researcharxiv-cs-cv
2 Jun 2026
Applications

Evaluating Bivariate Causal Statements Based on Mutual Compatibility

DGX agent

arXiv:2606.00278v1 Announce Type: new Abstract: For many real-world systems, causal ground truth is difficult to obtain, making claims about causal effects hard to assess. We develop methods for evalu

applicationsarxiv-cs-ai
2 Jun 2026
Tutorials

Evidence-Gated LLM Priors for Multi-Objective Bayesian Optimization

DGX agent

arXiv:2606.01730v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as heuristic advisors for black-box optimization, yet their suggestions and self-reported confidence

tutorialsarxiv-cs-ai
2 Jun 2026
Safety

Family Matters: A Systematic Study of Spatial vs. Frequency Masking for Continual Test-Time Adaptation

DGX agent

arXiv:2512.08048v3 Announce Type: replace Abstract: Recent continual test-time adaptation (CTTA) methods adopt masked image modeling to stabilize learning under distribution shift, yet each treats its

safetyarxiv-cs-cv
2 Jun 2026
Research

Fast and Lightweight Novel View Synthesis with Differentiable Multiplane Image

DGX agent

arXiv:2606.02068v1 Announce Type: cross Abstract: Recently, novel view synthesis has witnessed remarkable progress, with mainstream methods such as Neural Radiance Fields (NeRF) and 3D Gaussian Splatt

researcharxiv-cs-ai
2 Jun 2026
Tutorials

Finding What Matters: Anchoring Context Knowledge with Evolving Indices for Iterative Retrieval

DGX agent

arXiv:2601.16462v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) has become a dominant paradigm for mitigating hallucinations in Large Language Models (LLMs) by incorporating e

tutorialsarxiv-cs-cl
2 Jun 2026
Safety

Frequentist Consistency of Prior-Data Fitted Networks for Causal Inference

DGX agent

arXiv:2603.12037v2 Announce Type: replace Abstract: Foundation models based on prior-data fitted networks (PFNs) have shown strong empirical performance in causal inference by framing the task as an i

safetyarxiv-cs-lg
2 Jun 2026
Research

From Global to Local: Learning Context-Aware Graph Representations for Document Classification and Summarization

DGX agent

arXiv:2603.00021v2 Announce Type: replace Abstract: Recent NLP systems commonly represent documents as linear token sequences. Although this captures sequential order, it can hinder modeling long-rang

researcharxiv-cs-cl
2 Jun 2026
Research

From Layers to Submodules: Rethinking Granularity in Replacement-Based LLM Compression

DGX agent

arXiv:2606.02559v1 Announce Type: cross Abstract: Post-training compression of Large Language Models (LLMs) removes entire architectural components, either deleting them or replacing them with fitted

researcharxiv-cs-ai
2 Jun 2026
Applications

Graph-Augmented Retrieval for Cross-Entity Financial Sentiment Analysis: A Comparative Study

DGX agent

arXiv:2606.00062v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become foundational for grounding large language models in domain-specific corpora, yet conventional vector-bas

applicationsarxiv-cs-cl
2 Jun 2026
Research

Graph Transfer Learning via Shared Latent Geometry: Theory and Applications

DGX agent

arXiv:2606.00716v1 Announce Type: new Abstract: Inference and control in engineered physical systems pay a heavy physics cost at deployment: state estimators, inverse-problem solvers, model-predictive

researcharxiv-cs-lg
2 Jun 2026
Safety

Hard Labels In! Rethinking the Role of Hard Labels in Mitigating Local Semantic Drift

DGX agent

arXiv:2512.15647v3 Announce Type: replace Abstract: Soft labels from teacher models are a de facto practice for knowledge transfer and large-scale dataset distillation (e.g., SRe2L, LPLD). However, wh

safetyarxiv-cs-cv
2 Jun 2026
Safety

HASTE: Hardware-Aware Dynamic Sparse Training for Large Output Spaces

DGX agent

arXiv:2606.01117v1 Announce Type: cross Abstract: Extreme multi-label classification (XMC) involves learning models over large output spaces with millions of labels, making the output layer a memory-c

safetyarxiv-cs-ai
2 Jun 2026
Tools

Holo3.1: Fast & Local Computer Use Agents

DGX agent

Holo3.1 is a computer use agent developed by Hugging Face that enables fast, local execution of tasks on computing systems without requiring cloud infrastructure. The model is designed to interpret an

toolshugging-face
2 Jun 2026
Research

Honey, I Shrunk the Arc de Triomphe!

DGX agent

arXiv:2606.02379v1 Announce Type: new Abstract: Metric scale monocular geometry estimation has seen significant progress through large-scale data aggregation, yet current foundation models suffer from

researcharxiv-cs-cv
2 Jun 2026
Research

How Far Do Auto-Interpretation Labels Generalize: A Controlled Study Across Languages, Scripts, and Rewordings

DGX agent

arXiv:2606.00356v1 Announce Type: new Abstract: Sparse autoencoder (SAE) features are increasingly used to interpret language models, with auto-generated natural-language labels serving as the primary

researcharxiv-cs-cl
2 Jun 2026
Research

Hybrid Probabilistic Forecasting of Under-Five Malaria Admissions in Ghana: A Gaussian Process Regression with Holt-Winters Smoothing

DGX agent

arXiv:2606.00834v1 Announce Type: cross Abstract: Accurate malaria forecasting remains a major challenge in sub-Saharan Africa, where strong seasonality, reporting uncertainty, and non-stationary tran

researcharxiv-cs-ai
2 Jun 2026
Industry

If your daughter needs tutoring in algebra, you can probably find someone cheaper than Albert Einstein. Giving every task to GPT5.5 or Opus …

DGX agent

If your daughter needs tutoring in algebra, you can probably find someone cheaper than Albert Einstein. Giving every task to GPT5.5 or Opus 4.8 is overkill. Often times you can get the task done just

industryclem-delangue--x
2 Jun 2026
Research

Improving Visual Token Reduction via Rectifying Distortions for Efficient Multimodal LLM Inference

DGX agent

arXiv:2606.01711v1 Announce Type: new Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have achieved remarkable success in vision-language tasks, yet the quadratic computation

researcharxiv-cs-cv
2 Jun 2026
Research

Information-Theoretic Lower Bounds for Bit-Constrained Stochastic Optimization via a Reduction to Compressed Gaussian Mean Estimation

DGX agent

arXiv:2606.00703v1 Announce Type: cross Abstract: Low-precision pretraining (FP8, MXFP4, NVFP4) is now standard for frontier language models, yet the literature is almost entirely achievability -- alg

researcharxiv-cs-ai
2 Jun 2026
Safety

Interpretable Self-Supervised Learning via Representer Landmarks and Nystrom Approximation

DGX agent

arXiv:2509.24467v3 Announce Type: replace Abstract: Self-supervised learning (SSL) learns representations from massive unlabeled data, yet the resulting models typically operate as black boxes, necess

safetyarxiv-cs-lg
2 Jun 2026
Research

Interpreto: An Explainability Library for Transformers

DGX agent

arXiv:2512.09730v3 Announce Type: replace Abstract: Interpreto is an open-source Python library for interpreting HuggingFace language models, from early BERT variants to LLMs. It provides two compleme

researcharxiv-cs-cl
2 Jun 2026
Research

Is Zero-Shot Super-Resolution Possible in Operator Learning?

DGX agent

arXiv:2606.00296v1 Announce Type: cross Abstract: Neural operators are often reported to exhibit zero-shot super-resolution, a phenomenon in which a model trained on coarse grids produces accurate pre

researcharxiv-cs-lg
2 Jun 2026
Applications

It is difficult to know how good MAI-Thinking-1 is from the scores alone (like weirdly low GPQA & Terminal Bench 2.0) But Microsoft makes it…

DGX agent

It is difficult to know how good MAI-Thinking-1 is from the scores alone (like weirdly low GPQA & Terminal Bench 2.0) But Microsoft makes it really hard to try its models upon release (a general issue

applicationsethan-mollick--x
2 Jun 2026
Agents

Iteris: Agentic Research Loops for Computational Mathematics

DGX agent

arXiv:2606.02484v1 Announce Type: new Abstract: Recent advances in large language models and agentic AI systems have enabled significant progress in mathematical discovery, from solving competition pr

agentsarxiv-cs-ai
2 Jun 2026
Hardware

Join the livestream to hear from our team members @karan4d and @yoniebans live from @nvidia! https://www.youtube.com/watch?v=pgQDbRMa2Eg

DGX agent

Nous Research is hosting a livestream featuring team members Karan and Yoni speaking from NVIDIA, likely discussing AI research, model development, or collaboration between Nous Research and NVIDIA. T

hardwarenous-research--x
2 Jun 2026
Research

KACE: Knowledge-Adaptive Context Engineering for Mathematical Reasoning

DGX agent

arXiv:2606.00532v1 Announce Type: new Abstract: Context engineering can improve large language models without updating their weights, but mathematical reasoning exposes a key limitation: feedback accu

researcharxiv-cs-ai
2 Jun 2026
Tutorials

Last Layer Logits to Logic: Empowering LLMs with Logic-Consistent Structured Knowledge Reasoning

DGX agent

arXiv:2511.07910v2 Announce Type: replace Abstract: Large Language Models (LLMs) achieve excellent performance in natural language reasoning tasks through pre-training on vast unstructured text, enabl

tutorialsarxiv-cs-cl
2 Jun 2026
Agents

LeARN: Learnable and Adaptive Representations for Nonlinear Dynamics in System Identification

DGX agent

arXiv:2412.12036v2 Announce Type: replace Abstract: System identification, the process of deriving mathematical models of dynamical systems from observed input-output data, has undergone a paradigm sh

agentsarxiv-cs-lg
2 Jun 2026
Research

Learning Hamiltonian Dynamics at Scale: A Differential-Geometric Approach

DGX agent

arXiv:2509.24627v2 Announce Type: replace Abstract: Embedding physical intuition into network architectures allows the learning of dynamics that enforce fundamental properties, such as energy conserva

researcharxiv-cs-lg
2 Jun 2026
Safety

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning

DGX agent

arXiv:2606.02132v1 Announce Type: new Abstract: Agentic reinforcement learning can induce tool abuse, where models overuse external tools even for queries solvable by internal reasoning. Existing appr

safetyarxiv-cs-ai
2 Jun 2026
Safety

LLM Trainer: Automated Robotic Data Generation via Demonstration Augmentation using LLMs

DGX agent

arXiv:2509.20070v2 Announce Type: replace Abstract: We present LLM Trainer, a fully automated pipeline that leverages the world knowledge of Large Language Models (LLMs) to transform a small number of

safetyarxiv-cs-ro
2 Jun 2026
Tutorials

LLMs Need Encoders for Semantic IDs Too

DGX agent

arXiv:2606.00324v1 Announce Type: cross Abstract: Multimodal LLMs use dedicated encoders to bridge non-language modalities (vision encoders for images, depth models for audio codec tokens) because raw

tutorialsarxiv-cs-ai
2 Jun 2026
Hardware

Lodestar: An Online-Learning LLM Inference Router

DGX agent

arXiv:2606.00946v1 Announce Type: cross Abstract: Efficiently serving large language model (LLM) inference tasks is crucial both for user-perceived latency such as time-to-first-token (TTFT) and for G

hardwarearxiv-cs-ai
2 Jun 2026
Research

Machine Learning for Coding Retail Product Names to Consumer-Price Categories: A Rule-plus-Bag-of-Words Pipeline with Reliability-Weighted Human-in-the-Loop Labeling

DGX agent

arXiv:2606.02004v1 Announce Type: new Abstract: Consumer-price measurement increasingly draws on alternative data sources -- scanner, web-scraped, and transaction/receipt data. A recurring obstacle is

researcharxiv-cs-cl
2 Jun 2026
Safety

Measurement Geometry and Design for Trustworthy Generative Inverse Problems

DGX agent

arXiv:2606.02309v1 Announce Type: cross Abstract: Generative models are increasingly used as priors for inverse problems, but their ability to produce realistic images creates a basic trust problem: a

safetyarxiv-cs-cv
2 Jun 2026
Safety

MESA: Improving MoE Safety Alignment via Decentralized Expertise

DGX agent

arXiv:2606.00651v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) architectures scale Large Language Models (LLMs) efficiently, enabling greater capacity with reduced computational cost by dy

safetyarxiv-cs-ai
2 Jun 2026
Safety

Minimax-Optimal Policy Regret in Partially Observable Markov Games

DGX agent

arXiv:2606.02363v1 Announce Type: new Abstract: We study sequential decision-making in partially observable environments against strategic, adaptive opponents, modeled as partially observable Markov g

safetyarxiv-cs-lg
2 Jun 2026
Research

MINTS: Minimalist Thompson Sampling

DGX agent

arXiv:2606.01655v1 Announce Type: cross Abstract: The Bayesian paradigm offers principled tools for sequential decision-making under uncertainty, but its reliance on a probabilistic model for all para

researcharxiv-cs-ai
2 Jun 2026
Research

Normality-Preserving Continual Industrial Anomaly Detection via Orthogonal LoRA Banks

DGX agent

arXiv:2606.02042v1 Announce Type: new Abstract: Continual industrial anomaly detection with diffusion models suffers from historical normality prior drift and catastrophic forgetting. Existing continu

researcharxiv-cs-cv
2 Jun 2026
Agents

Not All Flips Are Conformity: Decomposing Stance Convergence in Multi-Agent LLM Debate

DGX agent

arXiv:2606.00820v1 Announce Type: new Abstract: Multi-agent debate (MAD) is a promising strategy for improving LLM reasoning, but when agents converge on a shared answer, it is unclear whether that co

agentsarxiv-cs-cl
2 Jun 2026
Tutorials

Not All Points Are Equal: Uncertainty-Aware 4D LiDAR Scene Synthesis

DGX agent

arXiv:2606.02510v1 Announce Type: new Abstract: Constructing faithful 4D worlds from LiDAR-acquired sequences is crucial for embodied AI, yet current generative frameworks apply uniform modeling capac

tutorialsarxiv-cs-cv
2 Jun 2026
Research

Not What, But How: A Communicative Audit of LLM Response Framing

DGX agent

arXiv:2606.02493v1 Announce Type: new Abstract: Large language models (LLMs) are being increasingly used to answer subjective, information-seeking questions, where users are sensitive to how responses

researcharxiv-cs-cl
2 Jun 2026
Local Ai

Ollama can't list this C# game script, because it might cause destruction.

DGX agent

A Reddit post discussing an issue where Ollama (an AI model tool) refuses to process or list a C# game script due to safety concerns about potential destructive code. The post likely explores the limi

local-air-ollama
2 Jun 2026
← Previous
1…962963964965966…1292
Next →