3 ms·
Does anyone have a good way for doing high-stakes deep research across a number of models? i.e. send to OpenAI, Anthropic, Gemini then evaluate (perhaps LLM as
by monkeydust 1y ago
Does anyone have a good way for doing high-stakes deep research across a number of models? i.e. send to OpenAI, Anthropic, Gemini then evaluate (perhaps LLM as judge)? Does that yield some performance uplift or make it worse?