Raven 0.2.0 — The Harness of Harnesses, built for RSI. 🐦⬛
One harness can't be best at everything. Raven combines its own specialist harnesses (Research, Code, Design, Oncall) with the agents you already use (Claude Code, Codex and more) into one team.
And it's built for RSI, and not just at the skill level. The whole harness can be rewritten by AI: prompts, policies, strategy code, playbooks. Every sub-harness, including the orchestration layer itself, is its own instance that can be improved.
With Raven you can:
1. Orchestrate many agents as one team. Raven's sub-harnesses and external agents work in one task graph with shared memory across sub-agents, powered by leading orchestration (0.963 Node F1 on the Multi-Agent Orchestration Benchmark).
2. Run long, complex tasks. Oncall and proactive execution keep work going for days, from scientific research loops to shipping a full Godot game.
3. Build vertical agents with RSI. Use Raven's RSI to develop and refine an agent for your domain, and we'll optimize it with you. Experimental for now; reach out to the Raven team(Discord:
More in the video and slides below. Open source, Apache-2.0.
(lots of work made with Raven lives there, and much of this launch's material was made with Raven too)
显示更多
🚨 SlowMist TI Alert 🚨
MemTensor's AI memory tooling has been compromised: MemoryOS (PyPI), the company's open-source long-term memory library for LLM and AI agents, and memtensor/memos-cloud-openclaw-plugin (npm), the official plugin connecting it to the OpenClaw agent runtime.
Affected versions bundle cross-platform Go binaries that execute when the package is loaded or imported: MemoryOS==2.0.34 on PyPI, and plugin versions 0.1.21, 0.1.23 and 0.1.25 on npm. You are affected if the PyPI version has been imported in your environment, or if the npm plugin is installed and the OpenClaw gateway has been started.
Potential attacker actions include harvesting npm/PyPI tokens, GitHub/GitLab credentials, AWS keys, SSH keys, API tokens, environment secrets, and other developer credentials, with data sent to infrastructure under skyleen[.]fr. The affected npm plugin may also expose user prompt content.
Users should remove or downgrade affected packages to known-good versions (0.1.20 for npm and 2.0.33 for PyPI), terminate sckit processes, block associated infrastructure, review network activity, and rotate credentials accessible from affected environments.
You can also visit to check for free whether the npm packages, pip packages, domains, or IPs you use are safe.
Reference:
As always, stay vigilant!
显示更多
I built StoryComet with Fable 5.1 in three days. 15 animated children's books in five languages, with translated voiceovers, covers and UI.
Parents can track progress. Kids can take quizzes, collect pins and 3D souvenirs, and switch languages on the fly. Six books teach English, Spanish, French, Japanese or Chinese.
My three-day workflow:
- Day 1: prototype (and post on X)
- Day 2: 9 more books, languages, progress, quizzes, pins and souvenirs
- Day 3: polish, domain name, supabase, accounts, Stripe and optimization
The first three books are 100% free so you can try the whole experience. Members unlocks the other 12 at early-bird pricing.
The 3D models and tokens cost me hundreds of dollars, so your support means a lot.
显示更多
Our third-party e-mail provider has been breached. Please be aware that the email named ‘Critical Security Alert: STM32 Entropy Vulnerability’ is not coming from us, and it’s a phishing attempt. Do not click on any link.
We have taken down the domain, and we are investigating the situation, including how the hackers got access to our legit domain.
显示更多
everyone calling Reddit dead because ChatGPT stopped citing it is wrong
Promptwatch said Reddit's share of ChatGPT citations went from 3.83% to 0.52%
when in April it was the most cited domain in ChatGPT
the panic is fair
if you built your AI visibility plan on Reddit, you'd be worried about it going to waste
but almost nobody opened the other dataset
Ahrefs looked at 145M US Google results
when the "discussions and forums" block shows up, Reddit is still in it about 84% of the time
Reddit lost one channel (maybe, it's not definite yet), but it's still showing up in the other one your buyer uses before they've decided anything
and nobody has a comparable baseline for the ChatGPT number either:
- Promptwatch had Reddit at 3.83% of citations before the drop
- Ahrefs, same platform, around the same time, had it at 16.7%
- Semrush once put it near 60%
yet everyone quoted an 86% collapse off a number the industry measures four different ways
and nobody outside OpenAI can actually explain either one. we all rebuilt our strategy around a number one tracker caught
Reddit is miserable to market on
it takes months of being useful before anyone lets you mention what you built
so continue doing that
显示更多
Financial work depends on trustworthy sources, consistent definitions, accurate calculations and auditable outputs.
Introducing Ling-3.0-flash-Fin, a finance-enhanced version of Ling-3.0-flash, developed with financial institutions and domain experts.
With 124B total and 5.1B active parameters, it supports information retrieval, research, valuation modeling and report preparation across long reports, research materials and complex workbooks.
The model showed competitive results across FinFIRST, FinSearchComp Verified, FinCRAFT, FinanceAgent v1.1/v2, APEX-Agents, SpreadsheetBench v1/v2 and τ³-Banking.
We will open-source the model weights next week.
显示更多
With frontier models getting more capable by the month 🚀, everyone is talking about Recursive Self-Improvement (RSI). But how do we actually measure this progress? 🤔
Excited to present RSI-Exam — a benchmark testing whether AI agents can improve an existing method through autonomous, long-horizon experimentation, and, importantly, whether those improvements generalize to hidden data.
Across 88 executable research tasks spanning 6 domains, the trend is clear: Claude and GPT form the first tier, with a clear gap over the rest of the field. Yet there is still huge room for improvement.
More importantly, the trajectories tell both stories: agents that discover fundamentally better methods — and agents that spend hours rigorously optimizing the wrong idea.
Check it out:
显示更多
JUST IN: 🇨🇳 China political advisory body CPPCC member announces he "believes Bitcoin" can attract lots of wealth and innovators to Hong Kong 👀
"I believe that Bitcoin is one of the really good synergies...to attract more funds in our domains."
显示更多
just sold for $125,000.
The seller ran one appraisal; it auto-built the listing (title, description, strengths). A buyer typed the domain, landed on the page, and negotiated the deal from there.
Your appraisal is your sales page. Full case study:
显示更多
Grok Bot is getting ridiculous.
People are controlling robots, shipping games, buying domains + running entire repos from their phone.
10 wild examples: