AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
Applications

H-Probes: Extracting Hierarchical Structures From Latent Representations of Language Models

DGX agent

arXiv:2605.00847v1 Announce Type: new Abstract: Representing and navigating hierarchy is a fundamental primitive of reasoning. Large language models have demonstrated proficiency in a wide variety of

applicationsarxiv-cs-cl
5 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

How Prompts Move Language Model Behavior: Frames, Salience, and Construal as Semantic Control

DGX agent

arXiv:2512.12688v3 Announce Type: replace-cross Abstract: Prompt engineering is widely used to shape large language model behavior, yet it is often treated as a practical heuristic rather than as a fo

researcharxiv-cs-cl
5 May 2026
Research

InfoLaw: Information Scaling Laws for Large Language Models with Quality-Weighted Mixture Data and Repetition

DGX agent

arXiv:2605.02364v1 Announce Type: new Abstract: Upweighting high-quality data in LLM pretraining often improves performance, but in datalimited regimes, especially under overtraining, stronger upweigh

researcharxiv-cs-cl
5 May 2026
Model Releases

LLM-Foraging: Large Language Models for Decentralized Swarm Robot Foraging

DGX agent

arXiv:2605.01461v1 Announce Type: new Abstract: Swarm foraging algorithms, such as the central-place foraging algorithm (CPFA), typically rely on offline parameter optimization using genetic algorithm

model-releasesarxiv-cs-ro
5 May 2026
Safety

Model-Based Proactive Cost Generation for Learning Safe Policies Offline with Limited Violation Data

DGX agent

arXiv:2605.01356v1 Announce Type: new Abstract: Learning constraint-satisfying policies from offline data without risky online interaction is crucial for safety-critical decision making. Conventional

safetyarxiv-cs-lg
5 May 2026
Research

Noise is All You Need: Solving Linear Inverse Problems by Noise Combination Sampling with Diffusion Models

DGX agent

arXiv:2510.23633v2 Announce Type: replace-cross Abstract: Pretrained diffusion models have demonstrated strong capabilities in zero-shot inverse problem solving by incorporating observation informatio

researcharxiv-cs-cv
5 May 2026
Agents

open-weight LLMs have come a long way on agent tasks! but the harness you wrap them in matters just as much as the model itself, and arguabl…

DGX agent

open-weight LLMs have come a long way on agent tasks! but the harness you wrap them in matters just as much as the model itself, and arguably the interface you use to drive that harness matters even m

agentsharrison-chase--x
5 May 2026
Model Releases

OpenAI GPT-5 System Card

DGX agent

arXiv:2601.03267v2 Announce Type: replace Abstract: This is the system card published alongside the OpenAI GPT-5 launch, August 2025. GPT-5 is a unified system with a smart and fast model that answers

model-releasesarxiv-cs-cl
5 May 2026
Agents

PACE: Post-Causal Entropy Modeling for Learned LiDAR Point Cloud Compression

DGX agent

arXiv:2605.01320v1 Announce Type: new Abstract: LiDAR point cloud compression is vital for autonomous systems to handle massive data from high-resolution sensors. While learned entropy modeling built

agentsarxiv-cs-cv
5 May 2026
Local Ai

Parllama -- a terminal UI for Ollama model management and multi-provider LLM chat

DGX agent

Parllama is a TUI (Text UI) application designed for easy management and use of Ollama-based LLMs that also works with major cloud-provided LLMs. It provides core model management features including f

local-air-ollama
5 May 2026
Model Releases

PC-MNet: Dual-Level Congruity Modeling for Multimodal Sarcasm Detection via Polarity-Modulated Attention

DGX agent

arXiv:2605.02447v1 Announce Type: new Abstract: Multimodal sarcasm detection, which aims to precisely identify pragmatic incongruities between literal text and nonverbal cues, has gained substantial a

model-releasesarxiv-cs-cl
5 May 2026
Research

PubMed-Ophtha: An open resource for training ophthalmology vision-language models on scientific literature

DGX agent

arXiv:2605.02720v1 Announce Type: cross Abstract: Vision-language models hold considerable promise for ophthalmology, but their development depends on large-scale, high-quality image-text datasets tha

researcharxiv-cs-cl
5 May 2026
Safety

Rethinking Model Selection in VLM Through the Lens of Gromov-Wasserstein Distance

DGX agent

arXiv:2605.01325v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have enhanced traditional LLMs with visual capabilities through the integration of vision encoders. While recent works hav

safetyarxiv-cs-cv
5 May 2026
Model Releases

Scaling Vision Transformers for Functional MRI with Flat Maps

DGX agent

arXiv:2510.13768v2 Announce Type: replace Abstract: We study the problem of training self-supervised foundation models for functional MRI. Our main contributions are: (1) we introduce a new model fami

model-releasesarxiv-cs-cv
5 May 2026
Research

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models

DGX agent

arXiv:2605.01853v1 Announce Type: new Abstract: Large reasoning models (LRMs) generate extended solutions, yet it remains unclear whether these traces reflect substantive internal computation or merel

researcharxiv-cs-cl
5 May 2026
Safety

SpectraDINO: Bridging the Spectral Gap in Vision Foundation Models via Lightweight Adapters

DGX agent

arXiv:2605.02258v1 Announce Type: new Abstract: Vision Foundation Models (VFMs) pretrained on large-scale RGB data have demonstrated remarkable representation quality, yet their applicability to multi

safetyarxiv-cs-cv
5 May 2026
Model Releases

Spoken Language Identification with Pre-trained Models and Margin Loss

DGX agent

arXiv:2605.01905v1 Announce Type: cross Abstract: For the speaker-controlled spoken language identification task proposed in the TidyLang Challenge 2026, this paper proposes a language identification

model-releasesarxiv-cs-cl
5 May 2026
Applications

Stochastic Modeling of Human-Machine Authentication Channels under Partial Information Leakage

DGX agent

arXiv:2605.02102v1 Announce Type: cross Abstract: Reliable and secure human-machine communication is fundamental to IoT and cyber-physical ecosystems, where smartphones and wearables commonly serve as

applicationsarxiv-cs-lg
5 May 2026
Model Releases

StressEval: Failure-Driven Dynamic Benchmarking for Knowledge-Intensive Reasoning in Large Language Models

DGX agent

arXiv:2605.01939v1 Announce Type: new Abstract: Static benchmarks for LLMs are increasingly compromised by contamination and overfitting especially on knowledge intensive reasoning tasks While recent

model-releasesarxiv-cs-cl
5 May 2026
Research

Stylistic Attribute Control in Latent Diffusion Models

DGX agent

arXiv:2605.02583v1 Announce Type: new Abstract: Text-to-image diffusion models have revolutionized image synthesis and editing, but precise control over stylistic attributes remains a challenge, often

researcharxiv-cs-cv
5 May 2026
Model Releases

VAnim: Rendering-Aware Sparse State Modeling for Structure-Preserving Vector Animation

DGX agent

arXiv:2605.01517v1 Announce Type: new Abstract: Scalable Vector Graphics (SVG) animation generation is pivotal for professional design due to their structural editability and resolution independence.

model-releasesarxiv-cs-cv
5 May 2026
Research

When To Adapt? Adapting the Model or Data in Federated Medical Imaging

DGX agent

arXiv:2605.00892v1 Announce Type: new Abstract: Federated learning enables collaborative model training across medical institutions without sharing raw data, but its performance is often limited by do

researcharxiv-cs-cv
5 May 2026
Research

Federated Weather Modeling on Sensor Data

DGX agent

arXiv:2605.00322v1 Announce Type: new Abstract: Federated weather modeling on sensor data is a distributed system underpinned by federated learning, enabling multiple sensor data sources, including gr

researcharxiv-cs-lg
4 May 2026
Safety

Intelligent Elastic Feature Fading: Enabling Model Retrain-Free Feature Efficiency Rollouts at Scale

DGX agent

arXiv:2605.00324v1 Announce Type: cross Abstract: Large-scale ranking systems depend on thousands of features derived from user behavior across multiple time horizons. Typically requires model retrain

safetyarxiv-cs-lg
4 May 2026
Research

It's Never Too Late: Noise Optimization for Collapse Recovery in Trained Diffusion Models

DGX agent

arXiv:2601.00090v2 Announce Type: replace Abstract: Contemporary text-to-image models exhibit a surprising degree of mode collapse, as can be seen when sampling several images given the same text prom

researcharxiv-cs-cv
4 May 2026
Safety

NEW paper from Sakana AI (ICLR 2026). A 7B Conductor model just hit SOTA on GPQA-Diamond and LiveCodeBench by orchestrating other LLMs inste…

DGX agent

NEW paper from Sakana AI (ICLR 2026). A 7B Conductor model just hit SOTA on GPQA-Diamond and LiveCodeBench by orchestrating other LLMs instead of solving problems itself. (great paper! bookmark it!) T

safetydair-ai--x
4 May 2026
Research

Reward Modeling from Natural Language Human Feedback

DGX agent

arXiv:2601.07349v3 Announce Type: replace Abstract: Reinforcement Learning with Verifiable reward (RLVR) on preference data has become the mainstream approach for training Generative Reward Models (GR

researcharxiv-cs-cl
4 May 2026
Tutorials

When Do Diffusion Models learn to Generate Multiple Objects?

DGX agent

arXiv:2605.00273v1 Announce Type: new Abstract: Text-to-image diffusion models achieve impressive visual fidelity, yet they remain unreliable in multi-object generation. Despite extensive empirical ev

tutorialsarxiv-cs-cv
4 May 2026
Local Ai

Best Local Vision-Language Models?

DGX agent

This discussion thread explores lightweight vision-language models that can run locally, including options like Llama 3.2 Vision, Qwen2.5-VL, and SmolVLM2, optimized for tasks like OCR and visual ques

local-air-stablediffusion
3 May 2026
Applications

Aligning Perception, Reasoning, Modeling and Interaction: A Survey on Physical AI

DGX agent

arXiv:2510.04978v5 Announce Type: replace Abstract: The rapid advancement of embodied intelligence and world models has intensified efforts to integrate physical laws into AI systems, yet physical per

applicationsarxiv-cs-ai
1 May 2026
Model Releases

Beyond Semantics: Measuring Fine-Grained Emotion Preservation in Small Language Model-Based Machine Translation

DGX agent

arXiv:2604.27920v1 Announce Type: cross Abstract: Preserving affective nuance remains a challenge in Machine Translation (MT), where semantic equivalence often takes precedence over emotional fidelity

model-releasesarxiv-cs-ai
1 May 2026
Safety

Cross-Lingual Sentiment Misalignment: Auditing Multilingual Language Models for Inversion Risk, Dialectal Representation, and Affective Stability

DGX agent

arXiv:2602.17469v2 Announce Type: replace Abstract: Recent advances in multilingual representation learning aim to bridge the performance gap between high- and low-resource languages, yet their abilit

safetyarxiv-cs-cl
1 May 2026
Model Releases

Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection

DGX agent

arXiv:2501.04066v2 Announce Type: replace Abstract: As a special type of multimedia data, Lithography Hotspot Detection (LHD) training often requires stronger privacy protection than conventional mult

model-releasesarxiv-cs-lg
1 May 2026
Agents

From Context to Skills: Can Language Models Learn from Context Skillfully?

DGX agent

arXiv:2604.27660v1 Announce Type: new Abstract: Many real-world tasks require language models (LMs) to reason over complex contexts that exceed their parametric knowledge. This calls for context learn

agentsarxiv-cs-ai
1 May 2026
Applications

GourNet: A CNN-Based Model for Mango Leaf Disease Detection

DGX agent

arXiv:2604.27764v1 Announce Type: new Abstract: Mango cultivation is crucial in the agricultural sector, significantly contributing to economic development and food security. However, diseases affecti

applicationsarxiv-cs-cv
1 May 2026
Applications

Green Physics-Informed Machine Learning Models For Structural Health Monitoring

DGX agent

arXiv:2604.27638v1 Announce Type: new Abstract: Machine learning continues to emerge as an important tool to be utilised within structural engineering and structural health monitoring, due to its abil

applicationsarxiv-cs-lg
1 May 2026
Applications

Improving Calibration in Test-Time Prompt Tuning for Vision-Language Models via Data-Free Flatness-Aware Prompt Pretraining

DGX agent

arXiv:2604.27715v1 Announce Type: new Abstract: Test-time prompt tuning (TPT) has emerged as a promising technique for enhancing the adaptability of vision-language models by optimizing textual prompt

applicationsarxiv-cs-cv
1 May 2026
Safety

Mitigating Selection Bias in Large Language Models via Permutation-Aware GRPO

DGX agent

arXiv:2603.21016v2 Announce Type: replace-cross Abstract: Large language models (LLMs) used for multiple-choice and pairwise evaluation tasks often exhibit selection bias due to non-semantic factors l

safetyarxiv-cs-ai
1 May 2026
Research

RopeDreamer: A Kinematic Recurrent State Space Model for Dynamics of Flexible Deformable Linear Objects

DGX agent

arXiv:2604.28161v1 Announce Type: new Abstract: The robotic manipulation of Deformable Linear Objects (DLOs) is a fundamental challenge due to the high-dimensional, non-linear dynamics of flexible str

researcharxiv-cs-ro
1 May 2026
Research

Static Program Slicing Using Language Models With Dataflow-Aware Pretraining and Constrained Decoding

DGX agent

arXiv:2604.26961v1 Announce Type: cross Abstract: Static program slicing is a fundamental software engineering technique for isolating code relevant to specific variables. While recent learning-based

researcharxiv-cs-ai
1 May 2026
Safety

The Effects of Visual Priming on Cooperative Behavior in Vision-Language Models

DGX agent

arXiv:2604.27953v1 Announce Type: new Abstract: As Vision-Language Models (VLMs) become increasingly integrated into decision-making systems, it is essential to understand how visual inputs influence

safetyarxiv-cs-ai
1 May 2026
Agents

Understanding Adversarial Transferability in Vision-Language Models for Autonomous Driving: A Cross-Architecture Analysis

DGX agent

arXiv:2604.27414v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used in autonomous driving because they combine visual perception with language-based reasoning, supporti

agentsarxiv-cs-cv
1 May 2026
Tutorials

VIPaint: Image Inpainting with Pre-Trained Diffusion Models via Variational Inference

DGX agent

arXiv:2411.18929v2 Announce Type: replace-cross Abstract: Diffusion probabilistic models learn to remove noise added during training, generating novel data (e.g., images) from Gaussian noise through s

tutorialsarxiv-cs-ai
1 May 2026
Industry

A cybersecurity harbinger: Oracle front-runs AI model threat with new customer security advisory

DGX agent

SiliconANGLE was able to review an Oracle Corp. security alert that went out to customers this week. We believe it was a direct response to Anthropic PBC’s new Mythos artificial intelligence model, an

industrysiliconangle
30 Apr 2026
Applications

Analysing Lightweight Large Language Models for Biomedical Named Entity Recognition on Diverse Ouput Formats

DGX agent

arXiv:2604.25920v1 Announce Type: cross Abstract: Despite their strong linguistic capabilities, Large Language Models (LLMs) are computationally demanding and require substantial resources for fine-tu

applicationsarxiv-cs-ai
30 Apr 2026
Research

AnimateAnyMesh++: A Flexible 4D Foundation Model for High-Fidelity Text-Driven Mesh Animation

DGX agent

arXiv:2604.26917v1 Announce Type: new Abstract: Recent advances in 4D content generation have attracted increasing attention, yet creating high-quality animated 3D models remains challenging due to th

researcharxiv-cs-cv
30 Apr 2026
Model Releases

Beyond Fixed Formulas: Data-Driven Linear Predictor for Efficient Diffusion Models

DGX agent

arXiv:2604.26365v1 Announce Type: new Abstract: To address the high sampling cost of Diffusion Transformers (DiTs), feature caching offers a training-free acceleration method. However, existing method

model-releasesarxiv-cs-cv
30 Apr 2026
Industry

Featherless.ai pulls in $20M to scale serverless hosting for open-source AI models

DGX agent

Featherless.ai Inc., a serverless inference platform startup that hosts open-source artificial intelligence models, today revealed it has raised 20 million in new funding to expand its global infrastr

industrysiliconangle
30 Apr 2026
← Previous
1…205206207208209…1271
Next →