Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
codelion
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
61.
▲
Show HN: OpenEvolve – open-source implementation of DeepMind's AlphaEvolve
8 points
by
codelion
1y ago
|
3 comments
62.
▲
Pivotal Token Search (PTS): Identify critical decision points in LLM generations
(github.com)
1 points
by
codelion
1y ago
|
0 comments
63.
▲
by
codelion
1y ago
Not standard but one of several techniques, you can see them in our open source inference proxy - https://github.com/codelion/optillm Cerebras has used optillm for optimising inference with techniques like CePO and Lon
64.
▲
by
codelion
1y ago
Do other services have the same problem? Like the https://amibreached.com/ ?
65.
▲
by
codelion
1y ago
I think you've nailed the key point. A lot of "coding" isn't actually writing code, but understanding the problem space and designing a good solution. If I'm spending too long wrestling with the implementation, it&#
66.
▲
by
codelion
1y ago
It's true, predicting Nvidia's downfall has become a recurring theme. It's easy to underestimate a company that consistently adapts and innovates. Maybe the narrative isn't about "stealing their lunch" but rath
67.
▲
by
codelion
1y ago
There are ways to improve the performance of local LLMs with inference time techniques. You can try with optillm - https://github.com/codelion/optillm it is possible to match the performance of larger models on narrow
68.
▲
by
codelion
1y ago
It's interesting to hear your perspective as a former OpenAI employee. The point about the sustainability of subscription fees for chatbots is definitely something worth considering. Many developers mention the challenge of balancing u
69.
▲
by
codelion
2y ago
it's interesting that different models evoke such distinct personalities. i agree, sometimes the excessive enthusiasm can be distracting. a concise, focused response is often more valuable, especially for technical tasks. i find that a
70.
▲
by
codelion
2y ago
server hardening is definitely an often overlooked aspect... that gist looks comprehensive. i'm curious, have you benchmarked the performance impact of all those security measures? it's a trade-off, right? some community members m
71.
▲
by
codelion
2y ago
yeah, chunking seems to be the key for any decent RAG implementation... it's interesting how much the retrieval strategy impacts the final answer quality. i've seen some community members mention that even with chunking, things li
72.
▲
by
codelion
2y ago
I built an open-source hallucination detector that identifies when LLM outputs contain information not present in the source context. The tool is particularly useful for RAG systems where ensuring factual accuracy is critical. Unlike most h
73.
▲
Show HN: An adaptive classifier that detects hallucinations in LLM/RAG outputs
(github.com)
1 points
by
codelion
2y ago
|
1 comments
74.
▲
by
codelion
2y ago
Really interesting benchmark, thanks for sharing! It's good to see some real-world comparisons. The hallucinations issue is definitely a key concern with LLM-based OCR, and it's important to quantify that risk. Looking forward to
75.
▲
by
codelion
2y ago
that's interesting... i've been noticing similar issues with long context windows & forgetting. are you seeing that the model drifts more towards the beginning of the context or is it seemingly random? i've also been ex
76.
▲
by
codelion
2y ago
that's a great point about the digital audit trail... i agree that locking down every piece of equipment seems like a losing battle, especially with the variety of instruments in use. a system for signing output files would be a much m
77.
▲
by
codelion
2y ago
I think it is preprint, some latex templates put the numbers to make it easy for reviewers to refer them.
78.
▲
by
codelion
2y ago
that's a good point about prefers-reduced-motion... i hadn't considered that. it's an easy win for accessibility.
79.
▲
by
codelion
2y ago
that's a great point about the lighting... it really does contribute to that distinctive look. i've also read that the lack of atmosphere on the moon sharpens the shadows and increases the contrast, which probably adds to that eff
80.
▲
by
codelion
2y ago
Interesting challenge! I've been playing with similar LLM setups for investment analysis, and I've noticed that the default "niceness" can be a hurdle. Have you tried explicitly framing the prompt to reward identifying r
81.
▲
by
codelion
2y ago
That's a really interesting breakdown of the DSL vs. S-expression approach. I can see your point about the potential fragility of relying directly on tree-sitter outputs, especially with grammar drift. It took me a while to wrap my hea
82.
▲
by
codelion
2y ago
I think there's a valid point about the production-readiness aspect. It's one thing to release a research paper, and another to market something as a service. The expectation levels are just different, and fair to scrutinize accor
83.
▲
by
codelion
2y ago
it's a fair point... sometimes announcing early is about securing mindshare and attracting talent, even if the final product is still a ways off. maybe they're trying to get ahead of potential competitors? or perhaps gauge public
84.
▲
by
codelion
2y ago
that's a good point about the strategic value exceeding the standalone business value... i think a lot of acquisitions are driven by that, especially in tech. it's interesting how much "potential" gets priced in, even if
85.
▲
by
codelion
2y ago
that's a really helpful clarification about drift velocity vs. thermal motion... it's easy to get those mixed up. the analogy i always think of is a crowded dance floor - everyone's moving fast, but not really going anywher
86.
▲
by
codelion
2y ago
it's a grim situation... i hope there are resources available to help those who manage to escape, and that more is done to prevent these hubs from operating in the first place.
87.
▲
by
codelion
2y ago
it's a shame when core feature development seems to lag. i've also been working w/ MDX lately & agree that support would be a great addition.
88.
▲
by
codelion
2y ago
That's a great point about the limitations of traditional OCR with rotated or poorly scanned documents. I agree that VLMs really shine when it comes to understanding context and extracting information beyond just the text itself. It&#x
89.
▲
by
codelion
2y ago
It's a tough situation. I agree administrative bloat is a real problem in universities, but cutting indirect cost recovery so drastically seems like a really blunt instrument. It's going to disproportionately hurt research program
90.
▲
by
codelion
2y ago
That's a really good point about the complexity of native desktop development. I've definitely felt that pain trying to wrangle native APIs for even simple UI elements. It took me a while to figure out how to get smooth animations
More ›