AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlog
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,103 results
Model Releases

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few…

DGX agent

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few hours building an LLM wiki with an agent powered entirely b

model-releasesdair-ai--x
1 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

ImagineNav++: Prompting Vision-Language Models as Embodied Navigator through Scene Imagination

DGX agent

arXiv:2512.17435v3 Announce Type: replace Abstract: Visual navigation is a fundamental capability for autonomous home-assistance robots, enabling long-horizon tasks such as object search. While recent

agentsarxiv-cs-ro
1 May 2026
Model Releases

KellyBench: A Benchmark for Long-Horizon Sequential Decision Making

DGX agent

arXiv:2604.27865v1 Announce Type: new Abstract: Language models are saturating benchmarks for procedural tasks with narrow objectives. But they are increasingly being deployed in long-horizon, non-sta

model-releasesarxiv-cs-ai
1 May 2026
Research

Qualitative Evaluation of Language Model Rescoring in Automatic Speech Recognition

DGX agent

arXiv:2604.27533v1 Announce Type: new Abstract: Evaluating automatic speech recognition (ASR) systems is a classical but difficult and still open problem, which often boils down to focusing only on th

researcharxiv-cs-cl
1 May 2026
Applications

RayFormer: Modeling Inter- and Intra-Ray Similarity for NeRF-Based Video Snapshot Compressive Imaging

DGX agent

arXiv:2604.27702v1 Announce Type: new Abstract: Video snapshot compressive imaging (SCI) enables the reconstruction of dynamic scenes from a single snapshot measurement. Recently, NeRF-based methods h

applicationsarxiv-cs-cv
1 May 2026
Research

Revealing the Impact of Visual Text Style on Attribute-based Descriptions Produced by Large Visual Language Models

DGX agent

arXiv:2604.27553v1 Announce Type: new Abstract: When the visual style of text is considered, a wide variety can be observed in font, color, and size. However, when a word is read, its meaning is indep

researcharxiv-cs-cv
1 May 2026
Industry

Standard Intelligence raises $75M to develop efficient computer use models

DGX agent

Standard Intelligence Inc., a six-person artificial intelligence startup, today announced that it has raised 75 million in funding. Sequoia and Spark Capital led the round. They were joined by multipl

industrysiliconangle
1 May 2026
Research

Exploring the Potential of Probabilistic Transformer for Time Series Modeling: A Report on the ST-PT Framework

DGX agent

arXiv:2604.26762v1 Announce Type: cross Abstract: The Probabilistic Transformer (PT) establishes that the Transformer's self-attention plus its feed-forward block is mathematically equivalent to Mean-

researcharxiv-cs-ai
30 Apr 2026
Model Releases

ext{PKS}^4:Parallel Kinematic Selective State Space Scanners for Efficient Video Understanding

DGX agent

arXiv:2604.26461v1 Announce Type: new Abstract: Temporal modeling remains a fundamental challenge in video understanding, particularly as sequence lengths scale. Traditional video models relying on de

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation

DGX agent

arXiv:2511.20714v2 Announce Type: replace-cross Abstract: World models serve as core simulators for fields such as agentic AI, embodied AI, and gaming, capable of generating long, physically realistic

model-releasesarxiv-cs-ai
30 Apr 2026
Research

Integrating Weather Foundation Model and Satellite to Enable Fine-Grained Solar Irradiance Forecasting

DGX agent

arXiv:2603.14845v3 Announce Type: replace-cross Abstract: Accurate day-ahead solar irradiance forecasting is essential for integrating solar energy into the power grid. However, it remains challenging

researcharxiv-cs-ai
30 Apr 2026
Model Releases

PATCH: Learnable Tile-level Hybrid Sparsity for LLMs

DGX agent

arXiv:2509.23410v4 Announce Type: replace-cross Abstract: Large language models (LLMs) deliver impressive performance but incur prohibitive memory and compute costs at deployment. Model pruning is an

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Preserving Disagreement: Architectural Heterogeneity and Coherence Validation in Multi-Agent Policy Simulation

DGX agent

arXiv:2604.26561v1 Announce Type: cross Abstract: Multi-agent deliberation systems using large language models (LLMs) are increasingly proposed for policy simulation, yet they suffer from artificial c

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior…

DGX agent

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior we want our agent to exhibit in production Model+Harness pu

model-releasesharrison-chase--x
30 Apr 2026
Model Releases

Training-Free Loosely Speculative Decoding: Accepting Semantically Correct Drafts Beyond Exact Match

DGX agent

arXiv:2511.22972v3 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance across diverse tasks but suffer from high inference latency due to their autoregressive gene

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

AdaTooler-V: Adaptive Tool-Use for Images and Videos

DGX agent

arXiv:2512.16918v3 Announce Type: replace Abstract: Recent advances have shown that multimodal large language models (MLLMs) benefit from multimodal interleaved chain-of-thought (CoT) with vision tool

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

DRAGON: A Benchmark for Evidence-Grounded Visual Reasoning over Diagrams

DGX agent

arXiv:2604.25231v1 Announce Type: cross Abstract: Diagram question answering (DQA) requires models to interpret structured visual representations such as charts, maps, infographics, circuit schematics

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity

DGX agent

arXiv:2604.24827v1 Announce Type: new Abstract: Closed-source frontier labs do not disclose parameter counts, and the standard alternative -- inference economics -- carries 2imes+ uncertainty from har

model-releasesarxiv-cs-lg
29 Apr 2026
Research

Large Language Models Are Effective Human Annotation Assistants, But Not Good Independent Annotators

DGX agent

arXiv:2503.06778v3 Announce Type: replace Abstract: Event annotation is important for identifying market changes, monitoring breaking news, and understanding sociological trends. Although expert annot

researcharxiv-cs-cl
29 Apr 2026
Hardware

Practical exposure correction via compensation

DGX agent

arXiv:2212.14245v2 Announce Type: replace Abstract: In computer vision, correcting the exposure level is a fundamental task for enhancing the visual quality of observations with inappropriate lightnes

hardwarearxiv-cs-cv
29 Apr 2026
Research

Towards interpretable AI with quantum annealing feature selection

DGX agent

arXiv:2604.25649v1 Announce Type: new Abstract: Deep learning models are used in critical applications, in which mistakes can have serious consequences. Therefore, it is crucial to understand how and

researcharxiv-cs-lg
29 Apr 2026
Research

Use of What-if Scenarios to Help Explain Artificial Intelligence Models for Neonatal Health

DGX agent

arXiv:2410.09635v2 Announce Type: replace Abstract: Early detection of intrapartum risks enables timely interventions to prevent or mitigate adverse labor outcomes such as cerebral palsy. However, acc

researcharxiv-cs-lg
29 Apr 2026
Research

VOYAGER: A Training Free Approach for Generating Diverse Datasets using LLMs

DGX agent

arXiv:2512.12072v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly being used to generate synthetic datasets for the evaluation and training of downstream models. Howeve

researcharxiv-cs-cl
29 Apr 2026
Research

A2DEPT: Large Language Model-Driven Automated Algorithm Design via Evolutionary Program Trees

DGX agent

arXiv:2604.24043v1 Announce Type: new Abstract: Designing heuristics for combinatorial optimization problems (COPs) is a fundamental yet challenging task that traditionally requires extensive domain e

researcharxiv-cs-ai
28 Apr 2026
Research

BiMol-Diff: A Unified Diffusion Framework for Molecular Generation and Captioning

DGX agent

arXiv:2604.24089v1 Announce Type: new Abstract: Bridging molecular structures and natural language is essential for controllable design. Autoregressive models struggle with long-range dependencies, wh

researcharxiv-cs-cl
28 Apr 2026
Safety

BVI-Mamba: Video Enhancement Using a Visual State-Space Model for Low-Light and Underwater Environments

DGX agent

arXiv:2604.23655v1 Announce Type: new Abstract: Videos captured in low-light and underwater conditions often suffer from distortions such as noise, low contrast, color imbalance, and blur. These issue

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Can Aha Moments Be Fake? Identifying True and Decorative Thinking Steps in Chain-of-Thought

DGX agent

arXiv:2510.24941v3 Announce Type: replace Abstract: Large language models can generate long chain-of-thought (CoT) reasoning, but it remains unclear whether the verbalized steps reflect the models' in

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Citation-Driven Multi-View Training for Patent Embeddings: QaECTER and Sophia-Bench

DGX agent

arXiv:2604.22897v1 Announce Type: cross Abstract: Patent retrieval underpins critical decisions in innovation, examination, and IP strategy, yet progress has been hampered by the absence of benchmarks

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Comparative Study of Weighted and Coupled Second- and Fourth-Order PDEs for Image Despeckling in Grayscale, Color, SAR, and Ultrasound

DGX agent

arXiv:2604.23612v1 Announce Type: new Abstract: Partial Differential Equation (PDE)-based approaches have gained significant attention in image despeckling due to their strong capability to preserve s

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

DVPO: Distributional Value Modeling-based Policy Optimization for LLM Post-Training

DGX agent

arXiv:2512.03847v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has shown strong performance in LLM post-training, but real-world deployment often involves noisy or incomplete su

safetyarxiv-cs-ai
28 Apr 2026
Agents

Enabling Transparent Cyber Threat Intelligence Combining Large Language Models and Domain Ontologies

DGX agent

arXiv:2509.00081v2 Announce Type: replace-cross Abstract: Effective Cyber Threat Intelligence (CTI) relies upon accurately structured and semantically enriched information extracted from cybersecurity

agentsarxiv-cs-ai
28 Apr 2026
Research

Generalizable Friction Coefficient Estimation via Material Embedding and Proxy Interaction Modeling

DGX agent

arXiv:2604.24188v1 Announce Type: new Abstract: Accurately estimating friction coefficients between arbitrary material pairs is critical for robotics, digital fabrication, and physics-based simulation

researcharxiv-cs-ro
28 Apr 2026
Applications

Gradient-Guided Exploration of Generative Model's Latent Space for Controlled Iris Image Augmentations

DGX agent

arXiv:2511.09749v2 Announce Type: replace Abstract: Developing reliable iris recognition and presentation attack detection methods requires diverse datasets that capture realistic variations in iris f

applicationsarxiv-cs-cv
28 Apr 2026
Model Releases

JudgeSense: A Benchmark for Prompt Sensitivity in LLM-as-a-Judge Systems

DGX agent

arXiv:2604.23478v1 Announce Type: new Abstract: Large language models are increasingly deployed as automated judges for evaluating other models, yet the stability of their verdicts under semantically

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines

DGX agent

arXiv:2604.23178v1 Announce Type: new Abstract: LLM-as-a-Judge has become the dominant paradigm for evaluating language model outputs, yet LLM judges exhibit systematic biases that compromise evaluati

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

MEASER: Malware embedding attacks on open-source LLMs

DGX agent

arXiv:2510.10486v2 Announce Type: replace-cross Abstract: Open-source large language models (LLMs) have demonstrated considerable dominance over proprietary LLMs in resolving neural processing tasks,

model-releasesarxiv-cs-ai
28 Apr 2026
Local Ai

Resource-Constrained UAV-Based Weed Detection for Site-Specific Management on Edge Devices

DGX agent

arXiv:2604.23442v1 Announce Type: new Abstract: Weeds compete with crops for light, water, and nutrients, reducing yield and crop quality. Efficient weed detection is essential for site-specific weed

local-aiarxiv-cs-cv
28 Apr 2026
Model Releases

Scheming Ability in LLM-to-LLM Strategic Interactions

DGX agent

arXiv:2510.12826v2 Announce Type: replace-cross Abstract: As large language model (LLM) agents are deployed autonomously in diverse contexts, evaluating their capacity for strategic deception becomes

model-releasesarxiv-cs-ai
28 Apr 2026
Research

SPEAR-1: Scaling Beyond Robot Demonstrations via 3D Understanding

DGX agent

arXiv:2511.17411v2 Announce Type: replace-cross Abstract: Robotic Foundation Models (RFMs) hold great promise as generalist, end-to-end systems for robot control. Yet their ability to generalize acros

researcharxiv-cs-lg
28 Apr 2026
Research

Adversarial Co-Evolution of Malware and Detection Models: A Bilevel Optimization Perspective

DGX agent

arXiv:2604.22569v1 Announce Type: cross Abstract: Machine learning-based malware detectors are increasingly vulnerable to adversarial examples. Traditional defenses, such as one-shot adversarial train

researcharxiv-cs-lg
27 Apr 2026
Model Releases

False Feasibility in Variable Impedance MPC for Legged Locomotion

DGX agent

arXiv:2604.22251v1 Announce Type: new Abstract: Variable impedance model predictive control (MPC) formulations that treat joint stiffness as an instantaneous decision variable operate on a feasible se

model-releasesarxiv-cs-ro
27 Apr 2026
Model Releases

How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals

DGX agent

arXiv:2604.22271v1 Announce Type: new Abstract: Large language models can detect their own errors and sometimes correct them without external feedback, but the underlying mechanisms remain unknown. We

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

How Popsa used Amazon Nova to inspire customers with personalised title suggestions

DGX agent

In this post, we share how we applied Amazon Bedrock and the Amazon Nova family of models to reimagine our Title Suggestion feature. By combining metadata, computer vision, and retrieval-augmented gen

model-releasesaws-ml-blog
27 Apr 2026
Model Releases

microsoft/VibeVoice

DGX agent

microsoft/VibeVoice VibeVoice is Microsoft's Whisper-style audio model for speech-to-text, MIT licensed and with speaker diarization built into the model. Microsoft released it on January 21st, 2026 b

model-releasessimon-willison
27 Apr 2026
Research

RouteLMT: Learned Sample Routing for Hybrid LLM Translation Deployment

DGX agent

arXiv:2604.22520v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved remarkable performance in Machine Translation (MT), but deploying them at scale remains prohibitively expensi

researcharxiv-cs-cl
27 Apr 2026
Research

Verbal Confidence Saturation in 3-9B Open-Weight Instruction-Tuned LLMs: A Pre-Registered Psychometric Validity Screen

DGX agent

arXiv:2604.22215v1 Announce Type: cross Abstract: Verbal confidence elicitation is widely used to extract uncertainty estimates from LLMs. We tested whether seven instruction-tuned open-weight models

researcharxiv-cs-ai
27 Apr 2026
Model Releases

Why there is no cloud version for Qwen 3.6 27/35B?

DGX agent

The Qwen 3.6-27B and 35B models are designed as open-weight models that developers can run locally on their own hardware without requiring cloud services. Alibaba released a separate cloud-only produc

model-releasesr-ollama
25 Apr 2026
Research

Attention-based multiple instance learning for predominant growth pattern prediction in lung adenocarcinoma wsi using foundation models

DGX agent

arXiv:2604.21530v1 Announce Type: cross Abstract: Lung adenocarcinoma (LUAD) grading depends on accurately identifying growth patterns, which are indicators of prognosis and can influence treatment de

researcharxiv-cs-ai
24 Apr 2026
← Previous
1…304305306307308…1294
Next →