How it feels to discover Hermes
Nous Research shared a post describing the experience of discovering or first encountering Hermes, their fine-tuned large language model series. The post likely captures the excitement and impressions
Knowledge catalogue
Nous Research shared a post describing the experience of discovering or first encountering Hermes, their fine-tuned large language model series. The post likely captures the excitement and impressions
I just love the language of this study...it speaks of the shifting 'community language'...and that is so true...have you noticed the new 'community language' is 'harness', it was 'contextual prompting
This Reddit thread from r/ChatGPT discusses user frustrations with ChatGPT's limitations in accessing credible and up-to-date information. Key issues raised include ChatGPT's training knowledge cutoff
Auto-generated synthesis of 142 entries about swyx--x
Harrison Chase, co-founder and CEO of LangChain, made a post on X affirming a point about where value lies in AI systems, specifically arguing that value is found 'in the harness' — referring to the s
Nous Research posted a weekend community engagement prompt encouraging developers and hobbyists to share what they are building using Hermes, their open-source large language model series. The post re
OpenAI offers an official ChatGPT desktop app for both macOS and Windows, available for free download at chatgpt.com/download. The app provides quick AI access via a keyboard shortcut (Option+Space...
This r/StableDiffusion thread discusses how to fine-tune the LTX-Video 2.3 model on a personal dataset, with the primary approach being LoRA (Low-Rank Adaptation), which fine-tunes a large AI model on
This r/MachineLearning post presents educational PyTorch implementations of FlashAttention versions 1 through 4, designed to highlight the key algorithmic differences across each iteration rather than
Nous Research posted a lighthearted social media message encouraging their community to take a break over the weekend from using Hermes, their AI model or platform. The post humorously acknowledges th
@hwchase17 Agent harnesess will simplify multi-agent orchestration and it is finally the next big thing, albeit it looks different from what we had imagined 3 year ago (AutoGen & CrewAI), 2 year ago (
Auto-generated index of all papers mentioned across the wiki.
Auto-generated index of all people mentioned across the wiki.
A Reddit post from the r/StableDiffusion community titled 'The classic UX you know and love' likely features a humorous or nostalgic reference to the user interface of a popular Stable Diffusion front
The specific Reddit post could not be retrieved from the search results. However, based on the context of the URL and related results, I can provide the following best-effort summary based on what ...
We recently identified a security issue involving the third-party developer library Axios that was part of a broader industry incident. We found no evidence that OpenAI user data was accessed, that ou
We're glad Hermes users have been making use of the free MiMo V2 Pro access via the Nous Portal! You loved it so much that we faced heavier initial usage than anticipated. Thank you to @Xiaomi for hel
arXiv:2604.07395v1 Announce Type: cross Abstract: Robotic manipulation systems that follow language instructions often execute grasp primitives in a largely single-shot manner: a model proposes an act
Leaders across industries around the world are asking: How do we harness all of this powerful technology effectively and at scale, to solve real problems, and drive value and impact, right now?Google
Agent skills are great. I wanted to share some of my favorites from Jesse Vincent's 'Superpowers' skill pack 𝚠𝚛𝚒𝚝𝚒𝚗𝚐-𝚙𝚕𝚊𝚗𝚜 skill produces much better plans than any harness' built in plan mode that I'
arXiv:2604.06296v1 Announce Type: cross Abstract: AI agents are increasingly deployed in real-world applications, including systems such as Manus, OpenClaw, and coding agents. Existing research has pr
arXiv:2604.07755v1 Announce Type: new Abstract: Despite extensive research, Large Language Models continue to hallucinate when generating code, particularly when using libraries. On NL-to-code benchma
**deepagents** (github.com/langchain-ai/deepagents) is an open-source, MIT-licensed agent harness built on LangChain and LangGraph, equipped with built-in task planning, a filesystem backend, subag...
arXiv:2604.06231v1 Announce Type: cross Abstract: Database systems incorporate an ever-growing number of functions in their kernels (a.k.a., database native functions) for scenarios like new applicati
arXiv:2604.05292v2 Announce Type: replace-cross Abstract: AI coding assistants are now used to generate production code in security-sensitive domains, yet the exploitability of their outputs remains u
arXiv:2604.07583v1 Announce Type: new Abstract: Real-world categorization is severely hampered by class imbalance because traditional ensembles favor majority classes, which lowers minority performanc
arXiv:2604.08523v1 Announce Type: new Abstract: AI agents may be able to automate your inbox, but can they automate other routine aspects of your life? Everyday online tasks offer a realistic yet unso
Completely agreed. Everyone in SF knows how good these models are at coding / work automation. When it comes to writing, there’s still a insane amount of model-generated AI slop because there’s no eas
I. We had crash-landed on the planet. We were far from home. The spaceship could not be repaired, and the rescue beacon had failed. Besides me, only the astrogator, part of the captain, and the ship’s
Cron Jobs — AI That Works on a Schedule Tell Qwen Code 'check if tests pass every 30 minutes' and it sets up a cron job in your session. No crontab editing, no scripts to write. Use /loop commend. Wor
The specific Reddit thread (r/MachineLearning post ID 1shg2ob) was not returned in the search results, and I was unable to directly fetch the URL's content. I cannot accurately summarize a page I h...
A father named Brady Frey discovered that his 12-year-old daughter had falsely registered her Discord account as 18+; he only found out after her account was hacked and he became trapped in a spir...
Every single one of these ads was made by one person. That’s insane, see the thread. Six months ago, that would have been impossible to say. The level of craft, story, and quality is quite impressive.
arXiv:2604.07070v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, their potential for purpose-driven exploration in dynamic geo-spatial
arXiv:2604.06219v1 Announce Type: cross Abstract: Across the Global North, calls for participatory artificial intelligence (AI) to improve the responsible, safe, and ethical use of AI have increased,
I was unable to retrieve the specific content of the linked tweet, as X (formerly Twitter) requires JavaScript and login to display individual posts. The search results did surface a highly relevan...
haha cool, this worked. an MCP server that allows Claude to build it's own reusable skills https://github.com/yoheinakajima/selfMCP (1473 LoC) basically a server with skills to CRUD skills in this vid
Nous Research draws its name from the ancient Greek concept of *nous* — the faculty of directly perceiving truth, reason, and divine reality — while its flagship model series, **Hermes**, is named ...
@IeloEmanuele is on a roll! you can now use claude code's advisor strategy with @LangChain agents because langchain/deepagents are provider agnostic, you can use different providers for your advisor a
arXiv:2604.07240v1 Announce Type: cross Abstract: We introduce a code-based challenge for automated, open-ended mathematical discovery based on the k-server conjecture, a central open problem in com
arXiv:2510.05524v2 Announce Type: replace Abstract: We present Knowledge Extraction on OMIn (KEO), a domain-specific knowledge extraction and reasoning framework with large language models (LLMs) in s
arXiv:2512.17445v2 Announce Type: replace Abstract: LangDriveCTRL is a natural-language-controllable framework for editing real-world driving videos to synthesize diverse traffic scenarios. It represe
A Reddit post in r/ChatGPT highlighted that ChatGPT does not natively have access to a real-time clock, meaning asking 'what time is it?' typically results in a disclaimer or an inaccurate response...
arXiv:2604.06846v1 Announce Type: cross Abstract: Interactive medical dialogue benchmarks have shown that LLM diagnostic accuracy degrades significantly when interacting with non-cooperative patients,
Microsoft is starting to remove 'unnecessary' Copilot buttons from its Windows 11 apps. In the latest version of the Notepad app for Windows Insiders, Microsoft has removed the Copilot button in favor
arXiv:2604.08516v1 Announce Type: new Abstract: Web agents--autonomous systems that navigate and execute tasks on the web on behalf of users--have the potential to transform how people interact with t
One great outcome of PaperWiki is personalized surveys. Survey papers continue to be one of the best ways to track a field. My agents are now generating personalized surveys on topics using my paper L
Open swe uses deepagents under the hood Deepagents is general purpose, openswe is focused on coding @hwchase17 @LangChain This is a very interesting comparison. Now my question is: how to compare Deep
arXiv:2604.07423v1 Announce Type: new Abstract: Physical Reservoir Computing (PRC) leverages the intrinsic nonlinear dynamics of physical substrates, mechanical, optical, spintronic, and beyond, as fi
The search results did not return the specific Reddit post about ibu-boost. Let me try fetching it directly. I was unable to retrieve the specific Reddit post or any direct information about the **...
TriNetX is a global health research network that connects pharmaceutical companies, study sites, investigators, and patients by leveraging real-world data (RWD) from electronic health records acros...
The Trump administration has ordered Reddit to appear before a secret grand jury in Washington, D.C., demanding the personal data of an anonymous user who criticized ICE online, with a compliance d...
arXiv:2604.07321v1 Announce Type: cross Abstract: Propositional Linear Temporal Logic (LTL) is a popular formalism for specifying desirable requirements and security and privacy policies for software,
The challenge with agent harnesses is not only building them but distribution. If we want agent products to scale like saas, the interfaces and ecosystem still need a fundamental breakthrough. im exci
In ChatGPT, a skill is a reusable, shareable workflow — defined by a `SKILL.md` file — that tells ChatGPT how to consistently execute a specific task without starting from scratch each time. Skills...
Sam Morrow, a Senior Software Engineer at GitHub on the Copilot Agent Services team, spoke at the ai.engineer conference about the challenges of building and scaling the GitHub MCP Server — an open...
A Reddit thread on r/ChatGPT discusses user frustration with ChatGPT resisting internet search requests and seemingly contradicting user-provided information. This behavior stems from two core issu...
Anthropic's new hosted service for long-running AI agents, designed to solve the challenge of creating systems that support 'programs as yet unthought of.' It abstracts infrastructure management to en
Editor’s note: This blog post outlines Google Cloud’s GPU AI/ML infrastructure reliability strategy, and will be updated with links to new community articles as they appear. As we enter the era of mul
Application security posture management company Apiiro Ltd. today announced the launch of a new command-line interface designed to bring application security directly into artificial intelligence-driv