AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Research

A Systematic Analysis of Hybrid Linear Attention

DGX agent

arXiv:2507.06457v2 Announce Type: replace Abstract: Transformers face quadratic complexity and memory issues with long sequences, prompting the adoption of linear attention mechanisms using fixed-size

researcharxiv-cs-cl
25 Jun 2026
Tools

[AINews] It's Meta-Harness Summer

DGX agent

Meta released Harness, an open-source framework for evaluating and benchmarking AI model performance across diverse tasks and datasets. The tool enables standardized testing of language models and aim

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
toolslatent-space
25 Jun 2026
Model Releases

An iterative energy-based multimodal transformer for joint retrieval of wheat soil moisture, leaf area index, and plant height from Sentinel-1 and Sentinel-2 time series

DGX agent

arXiv:2606.25174v1 Announce Type: cross Abstract: Field-scale retrieval of surface soil moisture (SM), leaf area index (LAI), and plant height (PH) is essential for precision agriculture, yet it remai

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Beyond One-Size-Fits-All: Diagnosis-Driven Online Reinforcement Learning with Offline Priors

DGX agent

arXiv:2606.25527v1 Announce Type: new Abstract: Online reinforcement learning (RL) agents increasingly depend on knowledge acquired offline to achieve practical efficiency. Originally studied in offli

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

CoLA: Cross-Modal Low-rank Adaptation for Multimodal Downstream Tasks

DGX agent

arXiv:2604.03314v2 Announce Type: replace-cross Abstract: Foundation models have revolutionized AI, but adapting them efficiently for multimodal tasks, particularly in dual-stream architectures compos

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Counterfeit Answers: Adversarial Forgery against OCR-Free Document Visual Question Answering

DGX agent

arXiv:2512.04554v2 Announce Type: replace Abstract: Document Visual Question Answering (DocVQA) enables end-to-end reasoning grounded on information present in a document input. While recent models ha

researcharxiv-cs-cv
25 Jun 2026
Model Releases

Distill on a Diet: Efficient Knowledge Distillation via Learnable Data Pruning

DGX agent

arXiv:2606.25488v1 Announce Type: new Abstract: Knowledge Distillation (KD) is widely used to obtain compact models for efficient inference in resource-constrained environments. Yet the computational

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Do Encoders Suffice? A Systematic Comparison of Encoder and Decoder Safety Judges for LLM Adversarial Evaluation

DGX agent

arXiv:2606.25782v1 Announce Type: new Abstract: With the widespread adoption of large language models (LLMs) in chatbots and everyday applications, companies increasingly need guardrails that are effe

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Generating Input Distributions for Explaining Portfolio Optimization Pipelines

DGX agent

arXiv:2606.25808v1 Announce Type: cross Abstract: We propose a predict-optimize-explain framework that uses gradient-based sample generation to interpret various portfolio models by identifying macroe

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

In a joint Fireworks and @Faros_AI evaluation of 211 real engineering tasks, Claude Code + GLM-5.2 beat both Claude Code + Opus 4.8 and Code…

DGX agent

In a joint Fireworks and @Faros_AI evaluation of 211 real engineering tasks, Claude Code + GLM-5.2 beat both Claude Code + Opus 4.8 and Codex + GPT-5.5: - Judge score: 0.568 vs. 0.521 and 0.466 - Time

model-releasesfireworks-ai--x
25 Jun 2026
Model Releases

Latent Block-Diffusion Temporal Point Processes: A Semi-Autoregressive Framework for Asynchronous Event Sequence Generation

DGX agent

arXiv:2606.24982v1 Announce Type: new Abstract: Modeling and sampling from the underlying distribution of asynchronous event sequences are crucial in various real-world applications, including social

model-releasesarxiv-cs-lg
25 Jun 2026
Safety

Learning Action Priors for Cross-embodiment Robot Manipulation

DGX agent

arXiv:2606.26095v1 Announce Type: cross Abstract: Most Vision-Language-Action (VLA) models build on a Vision-Language Model (VLM) backbone by attaching an action module and optimizing the full policy

safetyarxiv-cs-cv
25 Jun 2026
Model Releases

Memory-Efficient Policy Libraries with Low-Rank Adaptation in Reinforcement Learning

DGX agent

arXiv:2606.25700v1 Announce Type: new Abstract: When fine-tuning Large Language Models (LLMs), there has been success in minimizing both memory usage and computation with Parameter-Efficient Fine-Tuni

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

OpenAI staggers GPT-5.6 rollout for government vetting, eyes 2027 IPO

DGX agent

OpenAI Group PBC will roll out its next model, GPT-5.6, to a small group of partners rather than the public at the request of the Trump administration, the latest sign that Washington now wants to rev

model-releasessiliconangle
25 Jun 2026
Model Releases

PatchINR: Patch-Based Implicit Neural Representations for Efficient and Scalable Inference

DGX agent

arXiv:2606.25534v1 Announce Type: new Abstract: Implicit Neural Representation (INR) provides an effective approach for continuous signal modeling, but classical per-pixel inference results in quadrat

model-releasesarxiv-cs-cv
25 Jun 2026
Research

PERTINENCE: Input-based Opportunistic Neural Network Dynamic Execution

DGX agent

arXiv:2507.01695v3 Announce Type: replace Abstract: Deep neural networks (DNNs) are widely used for their ability to model complex patterns across domains such as computer vision, speech recognition,

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Project Auto-World: Towards Automated Benchmarking of Neural Relational Reasoners

DGX agent

arXiv:2606.24965v1 Announce Type: cross Abstract: Reasoning about relational structures remains a significant challenge for neural models, particularly when they must systematically apply learned know

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

RWGBench: Evaluating Scholarly Positioning in Related Work Generation

DGX agent

arXiv:2606.24894v1 Announce Type: cross Abstract: Large language models have shown strong fluency in scientific writing, yet the evaluation of related work generation (RWG) remains limited. Existing R

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Silent Failures in Physics-Informed Neural Networks: Parameter Poisoning and the Limits of Loss-Based Validation

DGX agent

arXiv:2606.25151v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) embed governing equations in their loss function, enabling mesh-free solutions to partial differential equation

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

The 4/elta Bound: Designing Predictable LLM-Verifier Systems for Formal Method Guarantee

DGX agent

arXiv:2512.02080v3 Announce Type: replace-cross Abstract: The integration of Formal Verification tools with Large Language Models (LLMs) offers a path to scale software verification beyond manual work

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

TopoCast: A Topological Fidelity Framework for Evaluating Transformer-Based Time Series Forecasting

DGX agent

arXiv:2606.25439v1 Announce Type: new Abstract: Deep learning-based models have achieved state-of-the-art performance in Time Series Forecasting (TSF), yet their evaluation remains dominated by pointw

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

TriViewBench: Controlled Complexity Scaling for Multi-View Structural Reasoning in MLLMs

DGX agent

arXiv:2606.26029v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) demonstrate strong performance on standard visual question answering benchmarks, yet their scalability under co

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Type Checking Project Haystack Grids using JSON Schema and Pydantic

DGX agent

arXiv:2606.24891v1 Announce Type: cross Abstract: Ontologies enable scalable energy services in buildings by supporting interoperability and automation. Project Haystack is a building ontology that is

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

When Multi-Sensor Fusion Fails to Generalize: Cattle Posture Classification Under Animal-Level and Temporal Distribution Shift

DGX agent

arXiv:2606.24986v1 Announce Type: new Abstract: Automated cattle posture-classification systems frequently report near-perfect accuracy, yet their robustness under realistic deployment conditions rema

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

3DCarGen: Scalable 3D Car Generation via 3D-consistent Multi-view Synthesis

DGX agent

arXiv:2606.24257v1 Announce Type: new Abstract: High-quality 3D vehicle assets are essential for autonomous driving simulation. Although multi-view diffusion-based paradigms enable controllable single

model-releasesarxiv-cs-cv
24 Jun 2026
Agents

AgentRivet: an automated system for producing Rivet routines from journal publications

DGX agent

arXiv:2606.13535v3 Announce Type: replace-cross Abstract: Particle physics collider experiments provide Rivet routines as part of the analysis preservation strategy for model-independent measurements.

agentsarxiv-cs-ai
24 Jun 2026
Agents

Beyond Bayer: Task-Optimal Sensor Co-Design for Robust Autonomous-Driving Segmentation

DGX agent

arXiv:2606.24096v1 Announce Type: cross Abstract: Robust perception underpins autonomous driving, and most recent progress comes from scaling the model-larger backbones, foundation models, and coopera

agentsarxiv-cs-ai
24 Jun 2026
Safety

Beyond Trajectory Imitation: Strategy-Guided Policy Optimization for LLM Reasoning

DGX agent

arXiv:2606.24064v1 Announce Type: new Abstract: Distilling reasoning capabilities from strong to weak language models typically involves imitating specific solution trajectories, effectively transferr

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

Big Tech's $1 trillion AI moat just got DESTROYED by a free Chinese download. Microsoft, Amazon, Google, and Meta are pouring fortunes into …

DGX agent

Big Tech's 1 trillion AI moat just got DESTROYED by a free Chinese download. Microsoft, Amazon, Google, and Meta are pouring fortunes into chips and data centers because they have been told that whoev

model-releasesgary-marcus--x
24 Jun 2026
Model Releases

btw Zai IPO'ed in Jan at HK$120 a share. when I first met @louszbd nobody really knew anyone using GLM's. now they have beat deepseek with t…

DGX agent

btw Zai IPO'ed in Jan at HK$120 a share. when I first met @louszbd nobody really knew anyone using GLM's. now they have beat deepseek with the world's undisputed top open model and in some respects (s

model-releasesswyx--x
24 Jun 2026
Industry

Building a Gitlawb Nodes adapter for Hugging Face buckets. Decentralized git storage should be able to live on the most supportive OSS infra…

DGX agent

This post discusses developing a GitLab nodes adapter to enable Hugging Face model storage on decentralized git infrastructure, leveraging open-source platforms for distributed model repository hostin

industryclem-delangue--x
24 Jun 2026
Model Releases

Decentralised AI Training and Inference with BlockTrain

DGX agent

arXiv:2606.24722v1 Announce Type: new Abstract: Frontier AI training is increasingly shaped by access to dense, centrally controlled accelerator clusters. This creates a structural advantage for hyper

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

How Much Can We Trust LLM Search Agents? Measuring Endorsement Vulnerability to Web Content Manipulation

DGX agent

arXiv:2606.16821v2 Announce Type: replace Abstract: Large language model (LLM)-based search agents synthesize open-web content into actionable recommendations on behalf of users, creating a risk that

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Jolia: Concept-Level Vision-Language Alignment for 3D CT Contrastive Learning

DGX agent

arXiv:2606.24570v1 Announce Type: new Abstract: Vision-language contrastive pretraining has become the dominant recipe for 3D medical foundation models, leveraging the large volumes of paired scans an

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Knowledge-Graph Grounding Helps LLMs Only for Out-of-Training Knowledge: A Controlled Study on Clinical Question Answering

DGX agent

arXiv:2606.22419v2 Announce Type: replace Abstract: A recent Nature Medicine study reports that general-purpose frontier LLMs outperform specialized retrieval-augmented clinical tools on medical bench

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Mind the Heads: Topological Representation Alignment for Multimodal LLMs

DGX agent

arXiv:2606.23885v1 Announce Type: cross Abstract: Representation alignment has emerged as an effective approach to improve Multimodal Large Language Models (MLLMs) by regularizing their internal repre

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Modality-Aware Out-of-Distribution Detection for Multi-Modal Action Recognition

DGX agent

arXiv:2606.24404v1 Announce Type: new Abstract: The incorporation of additional modalities into action recognition models increases their performance across a wide range of settings. However, how this

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

ParseBench is now also available on Papers with Code! Find it here: https://paperswithcode.co/benchmark/parsebench

DGX agent

ParseBench is now also available on Papers with Code! Find it here: https://paperswithcode.co/benchmark/parsebench We benchmarked Mistral OCR against other frontier and open-weight models on ParseBenc

model-releasesjerry-liu--x
24 Jun 2026
Model Releases

Point-Voxel Absorbing Graph Representation Learning for Event Stream based Recognition

DGX agent

arXiv:2306.05239v3 Announce Type: replace Abstract: Sampled point and voxel methods are usually employed to downsample the dense events into sparse ones. After that, one popular way is to leverage a g

model-releasesarxiv-cs-cv
24 Jun 2026
Research

Posterior Refinement: Fast Language Generation via Any-Order Flow Maps

DGX agent

arXiv:2606.24773v1 Announce Type: new Abstract: Non-autoregressive generation offers a powerful paradigm for iterative refinement, allowing models to recursively critique, erase and regenerate arbitra

researcharxiv-cs-cl
24 Jun 2026
Model Releases

Privacy-Preserving RAG via Multi-Agent Semantic Rewriting: Achieving Confidentiality Without Compromising Contextual Fidelity

DGX agent

arXiv:2606.24623v1 Announce Type: cross Abstract: Retrieval-Augmented Generation enhances large language models by incorporating external knowledge, but deploying it in sensitive scenarios risks priva

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Project Ariadne: Prompt-Conditioned Route Generation for Synthesis Planning

DGX agent

arXiv:2606.24184v1 Announce Type: new Abstract: Retrosynthetic planning seeks to connect a target molecule to commercially available starting materials through a multistep route. Classical planners co

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

Rule2Text: A Framework for Generating and Evaluating Natural Language Explanations of Knowledge Graph Rules

DGX agent

arXiv:2508.10971v2 Announce Type: replace-cross Abstract: Knowledge graphs (KGs) can be enhanced through rule mining; however, the resulting logical rules are often difficult for humans to interpret d

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

SignNet-1M: Large-Scale Multilingual Sign Language Video Dataset with Downstream Benchmarks

DGX agent

arXiv:2606.24361v1 Announce Type: new Abstract: Sign language models are typically trained on datasets captured under constrained conditions, with limited viewpoint, background, and signer-identity di

model-releasesarxiv-cs-cv
24 Jun 2026
Agents

Sol Video Inference Engine: Agent-Native Full-Stack Acceleration Framework for Efficient Video Generation

DGX agent

arXiv:2606.23743v1 Announce Type: cross Abstract: Modern video diffusion models achieve higher generation quality through scaling, but this also increases inference cost. Although many acceleration me

agentsarxiv-cs-ai
24 Jun 2026
Model Releases

The African Language Tax: Quantifying the Cost, Latency, and Context Penalty of Tokenizing African Languages in Frontier LLMs

DGX agent

arXiv:2606.24460v1 Announce Type: cross Abstract: Commercial large language models bill, scale latency, and budget context per token. Yet tokenizers assign more subword tokens to the same meaning in s

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

The Degeneracy Distillery

DGX agent

arXiv:2606.23838v1 Announce Type: new Abstract: When two or more parameters or labels produce similar data, they are degenerate, or hard to distinguish. Degeneracies render both label prediction and i

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

VisCritic: Visual State Comparison as Process Reward for GUI Agents

DGX agent

arXiv:2606.24525v1 Announce Type: new Abstract: GUI agents powered by vision-language models show strong potential for automating digital tasks, yet frequently fail in long-horizon scenarios due to th

model-releasesarxiv-cs-cv
24 Jun 2026
← Previous
1…451452453454455…1371
Next →