Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
chandureddyvari
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
by
chandureddyvari
1y ago
You can’t compare these with regular VM of aws or gcp. VM are expected to boot up in milliseconds and can be stopped/killed in milliseconds. You are charged per second of usage. The sandboxes are ephemeral and meant for AI coding agent
32.
▲
by
chandureddyvari
1y ago
For anyone curious about what LEAN is, like me, here’s the explanation: Lean Theorem Prover is a Microsoft project. You can find it here: https://www.microsoft.com/en-us/research/project/lean/
33.
▲
by
chandureddyvari
1y ago
There were other HN posts suggesting BMAD, ccpm, conductor, etc. I considered giving it a try. They were quite comprehensive, to the point where I was exhausted reading all the documentation they’ve generated before coding - product require
34.
▲
by
chandureddyvari
1y ago
Unsolicited advice: Why doesn’t open router provide hosting services for OSS models that guarantee non-quantised versions of the LLMs? Would be a win-win for everyone.
35.
▲
by
chandureddyvari
1y ago
I use Roo code with orchestrator(Boomerang) mode which pretty much has similar workflow. The orchestrator calls the architect to design the specs, and after iterating and agreeing on the approach, it is handed over to Code mode to execute t
36.
▲
by
chandureddyvari
1y ago
Getting rate limit error - This request would exceed the rate limit for your organization You should use something like openrouter or portkey or similar for managing fallbacks
37.
▲
by
chandureddyvari
1y ago
I’m exploring two different applications of AI for education and skill-building: 1. Open-Source AI Curriculum Generator(OSS MathAcademy alternative for other subjects) Think MathAcademy meets GitHub: an AI system that generates complete com
38.
▲
The latest Gemini 2.5 Pro reflects a 24-point Elo score jump on LMArena
(blog.google)
2 points
by
chandureddyvari
1y ago
|
0 comments
39.
▲
by
chandureddyvari
2y ago
Thanks just discovered https://ai.pydantic.dev/graph/ from the above link, which seems to use FSMs instead of DAGs.
40.
▲
by
chandureddyvari
2y ago
Neat idea! It would be helpful to have LLMs ranked from best to worst for a given GPU. Few other improvements I can think of: - Use natural language for telling offloading requirements. - Just year of the LLM launch of HF url can help if it
41.
▲
by
chandureddyvari
2y ago
Wow. Hate is unwarranted. Tirumala Anna Prasadam is pretty good. quoting from press release “By the normal standards in the present kitchen 12 huge vessels are steamed to produce rice at the rate of five persons for one kg rice. “But we ar
42.
▲
by
chandureddyvari
2y ago
I'm working on helping my wife get her print-on-demand Shopify store off the ground. She designs the products herself, but ran into challenges with SEO. So, I built a custom app that connects to her Shopify store via API, using GPT-4o-
43.
▲
by
chandureddyvari
2y ago
I'll try to slightly dive into Advaita philosophy here. Both the Upanishads and the Yoga Vasistha affirm that the distinction between the seer and the Self is an illusion created by the mind. When this illusion is dispelled, the onenes
44.
▲
by
chandureddyvari
2y ago
There’s a relevant quote from Swami Vivekananda on this From Karma Yoga, Chapter: I, Karma in its effect on character— What we say a man “knows”, should, in strict psychological language, be what he “discovers” or “unveils”; what a man “lea
45.
▲
by
chandureddyvari
2y ago
This article about dealing with recurrence in applications was an eye-opener for me. Definitely recommend giving it a read. https://github.com/bmoeskau/Extensible/blob/master/recurrenc...
46.
▲
by
chandureddyvari
3y ago
Also fungibility! I’m equally concerned about the specific “purpose“ and “time duration”, the Deputy T Governor gave an example to the press., quoting from the article: “Let’s say a school has given money to a student who won a prize to buy
47.
▲
The Reserve Bank of India (RBI) is introducing 'programmability' to the e-Rupee
(finshots.in)
3 points
by
chandureddyvari
3y ago
|
4 comments
48.
▲
by
chandureddyvari
3y ago
The CBDC(Central Bank Digital Currency) Retail (CBDC-R) pilot currently enables Person to Person (P2P) and Person to Merchant (P2M) transactions. It is now proposed to enable additional functionalities of programmability and offline capabil
49.
▲
by
chandureddyvari
3y ago
Sorry if this is a dumb question. Can someone explain why it’s called 8x7B(56B) but it has only 46.7B params? and it uses 12.9B params per token generation but there are 2 experts(2x7B) chosen by a 2B model? I’m finding it difficult to wrap
50.
▲
by
chandureddyvari
3y ago
I had lot of fun chatting with Pi. After some poking around to give it’s “system prompt”(long when it came back and prompt injection was a cool thing)., it said it was using some conversational frameworks like Grice’s principle etc. I tried
51.
▲
by
chandureddyvari
3y ago
Interesting. I thought anything >1million would need a vector db to scale on production. What was your machine config for running faiss? Also did you plan for redundancy or was it just faiss as a service VM?
52.
▲
by
chandureddyvari
3y ago
Yeah what you mentioned might be true. Currently our understanding on how LLMs really work behind the screens is limited. For example, there was a recent research[1] where LLM's accuracy is better if the context is added at the beginni
53.
▲
by
chandureddyvari
3y ago
The context length is limited, for gpt-3.5 it's 4k tokens, there are other offerings which offer upto 100k(claude). 100k tokens is ~1 book., but it priced steeply for each call. It's often wiser, cheaper to Retrieve the context fr
54.
▲
by
chandureddyvari
3y ago
slightly tangential, but where do people get awesome landing pages like linear( https://fig.io/ . has similar landing page) etc. Do they build them in-house or buy templates somewhere? Many of the recently launched YC compani
55.
▲
by
chandureddyvari
3y ago
yeah. my bad. Turning off relay, did show the right message. Thanks for that.
56.
▲
by
chandureddyvari
3y ago
It sounds great. But their banner is showing that my ip address is from Mumbai, whereas I’m actually in Bengaluru, India. That’s not really re-assuring. Maybe it’s just apple relay on my device that’s obfuscating my details. edit: my bad, h
57.
▲
by
chandureddyvari
3y ago
Interesting project. How does it fare on scaling? We were evaluating vector dbs and since we were into b2b saas, we are keen on the sharding, scaling and multi tenancy features. currently we are more inclined towards milvus. https:/&#
58.
▲
Achieve 98% cost reduction or 4% accuracy improvement over large LLM(GPT-4)
(arxiv.org)
2 points
by
chandureddyvari
3y ago
|
1 comments
59.
▲
by
chandureddyvari
3y ago
FrugalGPT is a simple yet flexible instantiation of LLM cascade which learns which combinations of LLMs to use for different queries in order to reduce cost and improve accuracy. Experiments show that FrugalGPT can match the performance of
60.
▲
by
chandureddyvari
3y ago
I’ve been using Screen zen for quite some time now. I found it to be incredibly useful. Curious to know how clearspace is different from screenzen, given the latter is free on app store. https://apps.apple.com/in/app&#
More ›