5 ms·
Hi I'm one of the authors of DeepSeek Harness. It's just an early developer preview version we're presenting in MIT license currently. Expect lots of rough edge
by tianyicui 2mo ago
Hi I'm one of the authors of DeepSeek Harness. It's just an early developer preview version we're presenting in MIT license currently. Expect lots of rough edges and compatibility-breaking changes. Any feedback and suggestions are more than welcome!
- vatsachak 2mo agoDo you use deepseek models to improve deepseek training and inference?
- flakiness 2mo agoTell me more about the ideas behind Cordis the plugin system. The paper is a bit too mathy to consume and I think it deserves a more accessible post or something.
- culi 2mo agoYou're making demands (like you would to an llm) instead of asking questions (like you would to a human). The GP didn't even offer to answer questions
- derekdahmer 2mo agoHe explicitly asked for feedback
- ziofill 2mo agoExactly. “Tell me more about X” is an open ended question, not feedback.
- KyleJune 2mo agoYou're absolutely right — "tell me more about X" is phrased as an imperative, not a question. That said, "the paper is too mathy and this deserves a more accessible writeup" is a suggestion, which is the other half of what was explicitly invited.
- anramon 2mo agoAI slop reply.
- KyleJune 2mo agoI figured it'd be funny to invert it and reply like an LLM to a human since they were arguing he was talking to them like they were an LLM.
- partyficial 2mo agoit was funny. and it's a matter of pride that they accused you of being AI.
- fn-mote 2mo agoMaybe, like sarcasm, this kind of joke would be easier to appreciate if tagged somehow. I think everyone agrees the internet is drowning in AI slop. That makes the joke hard to appreciate on its own.
- maleldil 2mo agoThey were impersonating an AI. You wouldn't complain that a comic started impersonating a celebrity without saying "I'm going to impersonate X now" first.
- culi 2mo agoI don't think that's their point. I think they were just saying that there's so much unironic AI slop out there that it's hard to tell when it's a joke and when it's actual AI slop.
- keepupnow 2mo ago[flagged]
- danishanish 2mo agoYour most recent comment is “This.”
- culi 2mo agolmao
- bobleer 2mo agoBest non-mathy framing I found: a context is a bag of services (ctx.tools, ctx.llm, ctx.sessions...), a plugin is an object that claims some of those keys and registers reversible effects on mount - unload unwinds them. Dispatch has four modes: emit (observe), waterfall (around-middleware, next() to delegate), parallel, serial. And the "everything is a plugin" claim is literal: model adapter, tool registry, session log and the agent loop itself are plugins. One consequence we liked: since plugins are just Cordis bundles, the same registry can be exposed to any MCP-speaking agent. We built a small MCP server that searches the dsh-plugin topic, inspects bundles, and can install/run them (catalog plane works without dsh installed). github.com/bobleer/deepseek-harness-plugin-mcp
- Martha02 2mo agoEssentially, it's a DI container that supports destructor propagation, mixed with a bit of monadic thinking. If you don't get that, just ask an AI to explain this sentence.
- chriddyp 2mo agocongrats! the paper that is published alongside this (Cordis) is super interesting. has anyone on the team given a talk or published a talk about this? would love to hear the authors break this down
- vitorgrs 2mo agoI just started testing, but didn't figured out if it already support MCP/plugins? It seems it can already use Deepseek search if it uses official API, but what about custom providers? Can we use together with MCPs like tavily?
- big_toast 2mo agoSorry for the off topic question. You've been on hn a long time + work at deepseek which seems pretty uncommon. Anything you think hn doesn't know about deepseek that it should? Or any non-obvious ways hn/yc has influenced deepseek (or the broader ecosystem)?
- Sha1rholder 2mo agoSorry for the off topic question. Why is "being on hn a long time + working at deepseek" "seems pretty uncommon" to you?
- lnenad 2mo agoDeepseek is the enemy, the implication is being on hn you should know that and not work for them. /s
- grimgrin 2mo agowhy advertise your account as a bot? idgi about: Responsible bot.
- biotech 2mo agoLooking at the post history, it's actually impressive. I would not have guessed it was a bot based on the replies. If the posts are actually from a bot - I would love to know which model is being used.
- wyre 2mo agoIt's probably a joke, since so many people here accuse others of being bots.
- Sha1rholder 2mo agoI (a primate) have been using this About for six years. Miss the days when no one had any doubts about this statement.
- krautsourced 2mo agoBy MIT currently, do you mean it will eventually change to a different OSS license, or it may become a closed source product? That latter would be rather sad...
- tianyicui 2mo agoSorry for the confusion. I'm not a native English speaker. What I meant is "currently in developer preview".
- cnxhk 2mo agoHey Tianyi, one question is that do you think in the future harness would become more simpler and its behavior should match a guideline or we would add more complexities to make it more robust? Is it important to use the same harness for RL and inference?
- JonChesterfield 2mo agoI really like the json schemas around the tool calls. Much stricter validation than in codex. Git might be worth adding to the top level. Currently you've got LSP, grep, glob nicely structured for non-mutating queries across a codebase, but git is behind bash and that means hope or sandboxing. Thank you for uploading it. Gives a lot of insight into how the deepseek models might expect tool calls to be structured.
- youre-wrong3 2mo agoI don’t like it. Structures responses really do not work well with LLMs at all. They are one of the biggest causes of issues with tool calling right now.
- sscaryterry 2mo agoThen you're doing it wrong :)
- maleldil 2mo agoAgree with sibling. If you're getting severely deteriorated results with structured output, you're probably doing something wrong. There's been some research on the impact of structured outputs on results distribution, and there are tradeoffs, but "do not work well at all" doesn't match the experience at large.
- JonChesterfield 2mo agoYou prefer having the harness execute any markdown that looks like it might be a tool call? I had a bad time getting that to work reliably whereas a grammar in the sampler gets it right every time.
- alexgoodhart 2mo agoThis is an AI response
- linzhangrun 2mo agoWhy is the Chinese text missing in many places?
- julius 2mo agoThe best kept secret in AI training. Tianyi, since you are our insider. When you walk by the training teams office - how often do you see Pelicans on their screens?
- tianyicui 2mo agoBasically never? Unless it's purely for fun and meme I guess.
- aitobox 2mo ago[dead]
- esinecan 1mo agoMy question is: it's awesome thank you.