Market Prices

BTC Bitcoin
$66,276.1 +1.59%
ETH Ethereum
$1,922.52 +1.31%
SOL Solana
$78.03 +0.46%
BNB BNB Chain
$573 +0.35%
XRP XRP Ledger
$1.14 +2.89%
DOGE Dogecoin
$0.0733 +1.90%
ADA Cardano
$0.1728 +2.13%
AVAX Avalanche
$6.55 -0.30%
DOT Polkadot
$0.8472 +2.88%
LINK Chainlink
$8.62 +0.87%

Event Calendar

{{年份}}
10
05
upgrade Ethereum Pectra Upgrade

Raises validator limit and account abstraction

18
03
unlock Sui Token Unlock

Team and early investor shares released

30
04
upgrade Celestia Mainnet Upgrade

Improves data availability sampling efficiency

15
04
halving Bitcoin Halving

Block reward reduced to 3.125 BTC

28
03
unlock Arbitrum Token Unlock

92 million ARB released

22
03
unlock Optimism Unlock

Circulating supply increases by about 2%

08
04
upgrade Solana Firedancer

Independent validator client goes live on mainnet

12
05
halving BCH Halving

Block reward halving event

Gas Tracker

Ethereum 28 Gwei
BNB Chain 3 Gwei
Polygon 42 Gwei
Arbitrum 0.5 Gwei
Optimism 0.3 Gwei

💡 Smart Money

0xde6e...f2cf
Market Maker
-$1.3M
73%
0x753f...0da0
Market Maker
+$3.9M
61%
0x2b16...13ca
Experienced On-chain Trader
+$4.3M
93%

🧮 Tools

All →

Perplexity Computer's WANDR: A Benchmark for AI Agents or Just Another Mirage?

0xIvy Industry
The announcement came through a whisper in the crypto media: Perplexity Computer, an entity linked to the AI search sensation Perplexity AI, has open-sourced a benchmark called WANDR for evaluating AI agents. On the surface, it sounds like a gift to the research community—a standardized way to measure how agents navigate, plan, and execute tasks. But in an industry where 'open-source' often masks a marketing agenda, the silence around the technical details screams louder than any press release. Perplexity AI built its reputation on a search engine that cites sources and answers questions with surprising accuracy. Now, it appears to be pivoting toward autonomous agents—programs that can act on users' behalf across websites and systems. WANDR, according to the sparse announcement, is supposed to accelerate AI research by providing a common ground for comparison. Yet, as someone who has spent years auditing blockchain protocols and their claims of 'decentralization,' I can't help but see parallels: too many projects promise standards but deliver only PowerPoint decks. The code compiles, but does it heal? That question echoes when I look at the history of benchmarks in both AI and crypto. In AI, we've seen benchmarks like GAIA and WebArena become de facto yardsticks, but they are often gamed or become obsolete as models improve. In crypto, we've witnessed countless 'liquidity benchmarks' designed to justify new tokens, only for them to fail under real market stress. Based on my audit experience, a benchmark's true value lies not in its name or the hype around its release, but in its transparency, reproducibility, and resistance to manipulation. WANDR—a name that suggests wandering or navigation—could be a meaningful addition if it measures what agents truly need: cross-platform coordination, error recovery, and ethical decision-making. But the announcement from Crypto Briefing is devoid of such details. No dataset size, no evaluation metrics, no baseline models, no comparison to existing work. It's like being handed a map with only a compass and no terrain. The silence is the loudest indicator of systemic rot—whether in a protocol or a research claim. I recall the Terra/Luna collapse in 2022. Before the crash, there was euphoria about algorithmic stability, creative tokenomics, and 'revolutionary' benchmarks of liquidity provision. Then the silence came—the silence of wallets zeroing out. That taught me to question every assertion that lacks public, auditable evidence. Similarly, WANDR may have been released in a bull market for AI agents, but that does not excuse the lack of substance. The market is euphoric, and FOMO is high; my job is to peel back the curtain. Let's consider the contrarian angle: even if WANDR is technically sound, the real challenge in AI agent evaluation is not the benchmark itself, but the incentive structures behind it. Benchmarks can be designed to favor specific models or architectures—a form of 'benchmark hacking' that gives an illusion of progress. In crypto, we saw this with Layer2 sequencers: decentralized in name, but often centralized in practice. Trust is not encrypted; it is woven through verifiable actions and community scrutiny. WANDR, if it remains opaque, risks being just another tool for self-promotion rather than genuine advancement. The feminist perspective in code teaches me to ask not 'how fast can it grow?' but 'who benefits and who is left behind?' AI agents will increasingly interact with financial systems, healthcare, and personal data. A benchmark that ignores safety, fairness, and accountability is not just incomplete—it's dangerous. WANDR's silence on these axes is troubling. I have argued for years that diversity in decision-making leads to more robust systems; the same applies to benchmarks. They must incorporate perspectives beyond performance metrics. In my own work founding a crypto education platform, I've seen how easy it is to mistake activity for progress. Projects ship code, but does it heal the trust deficits in our systems? WANDR could be a scaffold for trust if Perplexity Computer releases the full details—code, methodology, evaluation scripts—and invites external auditing. If they do, they will contribute to a desperately needed standard. If they don't, the open-source label is a veneer for a closed garden. The takeaway is forward-looking: we are at a moment where AI agents are poised to become the next interface for digital interaction—akin to how smartphones replaced desktops. The benchmarks we set now will shape research for years. WANDR could either be a lighthouse or a ghost ship. I will be watching its GitHub repository, the discussions in academic forums, and the eventual release of any associated model. Until then, the silence speaks louder than the pump. Let's not confuse code deployment with wisdom. Trust is not encrypted; it is woven, thread by thread, by those who dare to be transparent.

Fear & Greed

25

Extreme Fear

Market Sentiment

Altseason Index

43

Bitcoin Season

BTC Dominance Altseason

Market Cap

All →
# Coin Price
1
Bitcoin BTC
$66,276.1
1
Ethereum ETH
$1,922.52
1
Solana SOL
$78.03
1
BNB Chain BNB
$573
1
XRP Ledger XRP
$1.14
1
Dogecoin DOGE
$0.0733
1
Cardano ADA
$0.1728
1
Avalanche AVAX
$6.55
1
Polkadot DOT
$0.8472
1
Chainlink LINK
$8.62

🐋 Whale Tracker

🟢
0x94db...7c1a
1h ago
In
2,685 ETH
🔴
0xbe22...0fa2
30m ago
Out
3,540 ETH
🔴
0xfb23...8297
12h ago
Out
29,445 BNB