AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,569 results
Model Releases

Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routing

DGX agent

arXiv:2607.07953v1 Announce Type: cross Abstract: Self-attention lets each token retrieve information from the full context, but its quadratic cost in sequence length limits training and inference at

model-releasesarxiv-cs-ai
10 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

LiST: Lipschitz Scaling Training for Robust and Calibrated Neural Networks

DGX agent

arXiv:2607.07745v1 Announce Type: new Abstract: While accuracy, robustness, and calibration are all essential for reliable neural networks, they are often studied separately; developing models that sa

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

LlamaSeg: Image Segmentation via Autoregressive Mask Generation

DGX agent

arXiv:2505.19422v2 Announce Type: replace Abstract: We present extbf{LlamaSeg}, a visual autoregressive framework that unifies multiple image segmentation tasks via natural language instructions. By r

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

LSRM: High-Fidelity Object-Centric Reconstruction via Scaled Context Windows

DGX agent

arXiv:2604.05182v2 Announce Type: replace-cross Abstract: We introduce the Large Sparse Reconstruction Model to study how scaling transformer context windows affects feed-forward 3D reconstruction. Al

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

LUMI: Tokenizer-Agnostic LLM-Based Lossless Image Compression

DGX agent

arXiv:2607.08221v1 Announce Type: new Abstract: Large language model (LLM)-based lossless image compression methods typically represent pixel data through the native text interface of a pretrained mod

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

MetaHGNIE: Meta-Path Induced Hypergraph Contrastive Learning in Heterogeneous Knowledge Graphs

DGX agent

arXiv:2512.12477v2 Announce Type: replace Abstract: Estimating node importance in heterogeneous knowledge graphs is a fundamental problem underlying recommendation, search, and knowledge decision syst

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Mixture of Enhanced-View Experts for Multi-Query Vehicle ReID and A Large-Scale Benchmark

DGX agent

arXiv:2607.08085v1 Announce Type: new Abstract: Multi-query vehicle ReID aims to leverage complementary information from diverse views for robust feature learning. However, current methods suffer from

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

MSRNet: A Multi-Scale Recursive Network for Camouflaged Object Detection

DGX agent

arXiv:2511.12810v2 Announce Type: replace-cross Abstract: Camouflaged object detection is an emerging and challenging computer vision task that requires identifying and segmenting objects that blend s

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Multi-Resolution Feature Stem for Diabetic Retinopathy lesion segmentation

DGX agent

arXiv:2607.08679v1 Announce Type: new Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness worldwide, requiring automated lesion segmentation using deep learning models for

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

New research from Meta. (bookmark it) It's on how to fix agents that forget previously made decisions. It's well know that long-horizon agen…

DGX agent

New research from Meta. (bookmark it) It's on how to fix agents that forget previously made decisions. It's well know that long-horizon agents keep forgetting decisions they already made. Meta researc

model-releasesdair-ai--x
10 Jul 2026
Model Releases

@nickfrosst Full podcast > https://www.youtube.com/watch?v=pc4vT8xcSaY

DGX agent

Nick Frost, likely a notable figure in AI or machine learning, discusses his work and insights in a full-length podcast available on YouTube. The podcast was shared by Cohere, an AI company, suggestin

model-releasescohere--x
10 Jul 2026
Model Releases

Not everyone will succeed in the model business. 'You've seen a lot of companies spend a huge amount on compute and then not end up with a g…

DGX agent

Cohere warns that computational investment alone does not guarantee success in building AI models, emphasizing that many companies spend substantial resources on computing infrastructure without achie

model-releasescohere--x
10 Jul 2026
Model Releases

Okay, this is winning big time for me right now. Surprised how good GPT-5.6 is at verifiying/advising and all high-level orchestrator capabi…

DGX agent

This post discusses positive experiences with GPT-5.6's capabilities in verification, advisory functions, and high-level orchestration tasks, suggesting the model performs better than expected in thes

model-releasesdair-ai--x
10 Jul 2026
Model Releases

OmniFood-Bench: Evaluating VLMs for Nutrient Reasoning and Personalized Health Advice

DGX agent

arXiv:2607.08423v1 Announce Type: new Abstract: The rapid integration of Large Vision-Language Models (VLMs) into critical infrastructure promises to revolutionize personalized healthcare and dietary

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

On the Role of Conversational Timing in Synthetic Training Data for ASR

DGX agent

arXiv:2607.08371v1 Announce Type: cross Abstract: Synthetic multi-speaker conversations are widely used to train conversational automatic speech recognition (ASR) systems, but it remains unclear which

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

One of the most confusing aspects of GPT-5.6 is figuring out which model to use at which reasoning effort - sounds like Sol on Medium might …

DGX agent

One of the most confusing aspects of GPT-5.6 is figuring out which model to use at which reasoning effort - sounds like Sol on Medium might be a good new default for coding work, if it's an upgrade fr

model-releasessimon-willison--x
10 Jul 2026
Model Releases

OpenWiki general purpose memory is meant to be complementary to codex/claude code memory: it's proactive & ambient, meaning it'll automatica…

DGX agent

OpenWiki general purpose memory is meant to be complementary to codex/claude code memory: it's proactive & ambient, meaning it'll automatically go out into your world (via connections like gmail, x, n

model-releasesharrison-chase--x
10 Jul 2026
Model Releases

ParamMute: Suppressing Knowledge-Critical FFNs for Faithful Retrieval-Augmented Generation

DGX agent

arXiv:2502.15543v4 Announce Type: replace-cross Abstract: Large language models (LLMs) integrated with retrieval-augmented generation (RAG) have improved factuality by grounding outputs in external ev

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

path_boost: A Python Package for Interpretable Graph-Level Prediction using Path-Based Gradient Boosting

DGX agent

arXiv:2607.07935v1 Announce Type: cross Abstract: We present path_boost, a Python package for interpretable supervised learning on graph-structured input data. The package implements PathBoost, a grad

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Persistent Multiscale Density-based Clustering

DGX agent

arXiv:2512.16558v3 Announce Type: replace Abstract: Clustering is a cornerstone of modern data analysis. Detecting clusters in exploratory data analyses (EDA) requires algorithms that make few assumpt

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Persuasion Attacks Can Decrease Effectiveness of CoT Monitoring

DGX agent

arXiv:2607.08066v1 Announce Type: new Abstract: Chain-of-thought (CoT) monitoring is a promising safety mechanism for AI agents, based on the premise that visible reasoning traces can surface misalign

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

PhasorFlow: A Python Library for Unit Circle Based Computing

DGX agent

arXiv:2603.15886v3 Announce Type: replace-cross Abstract: We present PhasorFlow, an open-source Python library for computing on the S^1 unit circle. Inputs are encoded as complex phasors z=e^{iphi} on

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Physics-Informed Machine Learning Under Small-Data Constraints: Lessons from Abrasive Waterjet Milling

DGX agent

arXiv:2607.07863v1 Announce Type: new Abstract: In physically dominated machining processes, experimental datasets are small, expensive, and material-specific; in this regime, data curation, evaluatio

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

PolyUQuest: Verifiable Structure-Aware Web RAG over Heterogeneous Graphs

DGX agent

arXiv:2607.08269v1 Announce Type: new Abstract: Existing retrieval-augmented generation (RAG) systems treat web pages as flat text, losing the structural and semantic signals encoded in HTML. We prese

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Pose-to-Biomechanics: Bridging 3D Human Pose Estimation and Biomechanical Attribute Prediction

DGX agent

arXiv:2607.08725v1 Announce Type: cross Abstract: Recent progress in 3D human pose estimation has made markerless recovery of skeletal motion increasingly accurate and scalable. However, most pose est

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Psychological Competence as a Missing Dimension in AI Evaluation

DGX agent

arXiv:2607.08285v1 Announce Type: new Abstract: Current AI evaluation frameworks focus primarily on technical performance, including accuracy, robustness, reasoning ability, and policy compliance. The

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Real-World Blind Super-Resolution via Feature Matching with Implicit High-Resolution Priors

DGX agent

arXiv:2202.13142v3 Announce Type: replace Abstract: A key challenge of real-world image super-resolution (SR) is to recover the missing details in low-resolution (LR) images with complex unknown degra

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

ReCoLoRA: Spectrum-Aware Recursive Consolidation for Continual LLM Fine-Tuning

DGX agent

arXiv:2607.07719v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning adapts a large language model to one task cheaply, but across a task sequence LoRA-style methods keep stacking low-ran

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Reinforcing the Generation Order of Multimodal Masked Diffusion Models

DGX agent

arXiv:2607.08056v1 Announce Type: cross Abstract: Diffusion Language Models (DLMs) have recently achieved substantial progress in natural language generation tasks. Recent research demonstrates that a

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents

DGX agent

arXiv:2607.08716v1 Announce Type: new Abstract: In long-horizon tasks, decision-relevant state is often scattered across an expanding trajectory, while the action agent must surface it and act. As tra

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Resample or Reroute? Budget-Aware Test-Time Model Selection for Large Language Models

DGX agent

arXiv:2607.08665v1 Announce Type: new Abstract: Routing among large language models (LLMs) trades response quality against serving cost, motivated by the reported gap between deployed routers and a pe

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

RetailBench: Evaluating Long-Horizon Autonomous Decision-Making and Strategy Stability of LLM Agents in Realistic Retail Environments

DGX agent

arXiv:2603.16453v3 Announce Type: replace Abstract: Large language model (LLM) agents have made rapid progress on short-horizon, well-scoped tasks, yet their ability to sustain coherent decisions in d

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Scalable and Trustworthy Earth Observation Foundation Models

DGX agent

arXiv:2607.07758v1 Announce Type: new Abstract: Foundation models (FMs) have transformed machine learning from isolated task-specific model development toward general-purpose models pretrained on broa

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Secure Decentralized Federated Learning via Gossip and Virtual Voting

DGX agent

arXiv:2607.08651v1 Announce Type: new Abstract: Decentralized federated learning (DFL) removes the central server by letting nodes exchange model updates through peer-to-peer gossip, but existing goss

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Shift & Drift: A Zero-Shot Benchmark for Generalizable and Robust Autonomous Driving Motion Planning

DGX agent

arXiv:2607.07844v1 Announce Type: cross Abstract: While closed-loop motion planners trained on large-scale, object-level datasets, e.g., nuPlan, demonstrate strong in-distribution (ID) performance, th

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

SolarChain-Eval: A Physics-Constrained Benchmark for Trustworthy Economic Agents in Decentralized Energy Markets

DGX agent

arXiv:2607.08681v1 Announce Type: new Abstract: As agentic AI systems are increasingly applied to cyber-physical environments, their evaluation requires assessment of both task performance and trustwo

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

SQuaD-SQL: Efficient Text-to-SQL with Small Language Models via LLM-Guided Knowledge Distillation

DGX agent

arXiv:2607.08161v1 Announce Type: new Abstract: Text-to-SQL is a fundamental task in natural language processing that enables users to interact with structured databases using natural language. While

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Structured Pruning of Large Language Models via Power Transformation and Sign-Preserving Score Aggregation with Adaptive Feature Retention

DGX agent

arXiv:2607.08027v1 Announce Type: cross Abstract: This paper proposes an improved structured pruning method for large language models (LLMs) that addresses key challenges in adapting Adaptive Feature

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Super Weights in LLMs and the Failure of Selective Training

DGX agent

arXiv:2607.08733v1 Announce Type: new Abstract: Recent work identified Super Weights, individual parameters whose removal degrades model performance by orders of magnitude. We show that this degradati

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

SwinIFS: Landmark Guided Swin Transformer For Identity Preserving Face Super Resolution

DGX agent

arXiv:2601.01406v2 Announce Type: replace-cross Abstract: Face super-resolution aims to recover high-quality facial images from severely degraded low-resolution inputs, but remains challenging due to

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

TFP: Temporally Conditioned Memory-Fusion Policies for Visuomotor Learning

DGX agent

arXiv:2607.08283v1 Announce Type: new Abstract: Vision--Language--Action (VLA) policies such as pi_{0.5} and OpenVLA perform well on many manipulation tasks, but they are often reactive: the next acti

model-releasesarxiv-cs-ro
10 Jul 2026
Model Releases

The Download: Claude’s inner workings and OpenAI’s “super app”

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Anthropic found a hidden space where Claude puzzles over conce

model-releasesmit-tech-review
10 Jul 2026
Model Releases

The Phasor Transformer: Resolving Attention Bottlenecks on the Unit Circle

DGX agent

arXiv:2603.17433v2 Announce Type: replace-cross Abstract: Transformer models have redefined sequence learning, yet dot-product self-attention introduces a quadratic token-mixing bottleneck for long-co

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

The Regularization Parameter: Sparse Precision Matrix Estimation

DGX agent

arXiv:2607.07735v1 Announce Type: cross Abstract: Sparse precision matrix estimation provides an interpretable and computationally efficient framework for modeling conditional dependencies in high-dim

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

The same way, we're probably one of the few AI startups with user network effects, we might become the first one with agent network effects!

DGX agent

The same way, we're probably one of the few AI startups with user network effects, we might become the first one with agent network effects! Hugging Face Gemma Challenge results are in! 📈 Over 6 days,

model-releasesclem-delangue--x
10 Jul 2026
Model Releases

the sun is out today

DGX agent

the sun is out today To celebrate the launch of GPT-5.6 Sol, we will reset the rate limits again (twice) across ChatGPT Work and Codex over the next 24 hours. We want you to have the time to truly try

model-releasessam-altman--x
10 Jul 2026
Model Releases

This time it is novel math proofs with a public model (most of the other big math breakthroughs have been with experimental LLMs).

DGX agent

This time it is novel math proofs with a public model (most of the other big math breakthroughs have been with experimental LLMs). Yesterday, we made GPT-5.6 Sol Ultra generally available. Today, we'r

model-releasesethan-mollick--x
10 Jul 2026
Model Releases

This was a critical early paper on AI & work, showing that entrepreneurs getting advice from GPT-4 had higher profit margins if they were hi…

DGX agent

This was a critical early paper on AI & work, showing that entrepreneurs getting advice from GPT-4 had higher profit margins if they were high performing, but did worse if they were already in trouble

model-releasesethan-mollick--x
10 Jul 2026
← Previous
1…110111112113114…471
Next →