开发者分享 Jev 在定制工具中实现更高效工作流
Been integrating Jev into my custom harness. I couldn't be more excited about the results I am seei...
朋友在用 Jev 做定制工具,发现它比标准 LLM 更适合分类任务,能实现更便宜、更快的工作流,还解锁了动态 UI 和智能规划等新功能。
开发者正在将 Jev 集成到自己的定制工具中,发现它能实现更便宜和更快的工作流,帮助更好地扩展代理系统。Jev 还能通过智能决策和动态工作流管理提升可靠性,并解锁动态 UI 和更智能的代理规划等新功能。
Been integrating Jev into my custom harness. I couldn't be more excited about the results I am seei...
Been integrating Jev into my custom harness. I couldn't be more excited about the results I am seeing. Most demos on my timeline are flashy but very basic. Notice they are all about some boring classification task that was already possible before. I get it. Jev is fast and cheap. That point was made. But what new can Jev unlock? That's what I am more interested in: how does Jev evolve the agent harness experience? To seriously explore this, we first need to treat Jev as a key primitive, one of many to come. We must think about how it compliment today's test-time compute strategies. It's not competing; it's accelerating and allowing even more interesting scaling strategies. Jev is starting to unlock many insane and interesting experiences in my custom harnesses. Features that were either cost-prohibitive, lacked the right primitive, or were not feasible because of latency constraints. After a bit of exploration (excited to share more details soon), here are a few things that really excite me: - Jev enables cheaper & faster workflows (the obvious one), which can help us better scale our agent harnesses. If you need a classifier, want to improve control flow, or need a more deterministic workflow, don't use a standard LLM; Jev is probably a better fit. Jev is also great for labeling things at scale. - Jev improves reliability through components like intelligent decision-making, structured intelligence for dynamic workflows/generating harnesses on the fly, and new context management strategies; I've seen a few of the latter already, on potential ways to surface context on demand like tool calls, skill metadata, etc. So much to share here. - Jev unlocks new, interesting ways to enhance the agent experience (e.g., dynamic UIs, efficient LLM councils, smarter routing, enabling smarter, efficient orchestration and planning, more decisive, and useful proactive agents through surfacing more relevant and timely context). - Jev feels like the right primitive to enable more reliable evaluation strategies like LLM-as-a-Judge and building powerful verifiers for agent harnesses. These are two areas where Jev might unlock an insane amount of alpha. Not to mention, I see its potential to synthesize high-quality data to accelerate RSI by helping the model and harness co-evolve. I traveled this week (to Dreamforce) and am feeling extremely exhausted, but a full guide is dropping soon. 💬 3 🔄 2 ❤️ 10 👀 964 📊 5 ⚡