Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
vikp
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
61.
▲
by
vikp
3y ago
I think it's more about potential TAM than the current product. The hypothesis is probably that there will eventually be a winner in each vertical, and that $$ can mint the winner.
62.
▲
Show HN: Endless – Learn anything with personalized AI and blocks
(endless.academy)
1 points
by
vikp
3y ago
|
0 comments
63.
▲
by
vikp
3y ago
This would be interesting to me. There are a few options now, like Judge0, but the language versions are pretty out of date. Self-hosting is not a good time investment at the moment. Email me at hn at vikas.sh if you have a service. I&#x
64.
▲
by
vikp
3y ago
You mean "Multiply the vectors by the other vectors. This is attention - it's the magic of transformers, that enables combining information from multiple tokens together. This generates a new matrix."? It's really over
65.
▲
by
vikp
3y ago
Transformers are about converting some input data (usually text) to numeric representations, then modifying those representations through several layers to generate a target representation. In LLMs, this means go from prompt to answer. I&#
66.
▲
by
vikp
3y ago
It's unclear which models will be trained to 1.5T tokens. The details of how many tokens each model saw in training are on Github - https://github.com/stability-AI/stableLM/ . But only for the ones that hav
67.
▲
by
vikp
3y ago
It's fantastic that more orgs are releasing open-source models trained on more than 300B or so tokens. Here's my take from the details I could find. Pros - 4096 context width (vs 2048 for llama, gpt-j, etc) - 3B to 65B rele
68.
▲
by
vikp
3y ago
I've been making a course to help with this - https://github.com/VikParuchuri/zero_to_gpt . Lots of coding, with optional videos. It has very few prereqs.
69.
▲
by
vikp
3y ago
Attribution comes to mind.
70.
▲
by
vikp
4y ago
How does this compare to fauxpilot - https://github.com/fauxpilot/fauxpilot ? Fauxpilot also uses Triton with fastertransformers and GPT-J style models (codegen).
71.
▲
by
vikp
4y ago
If you want more open research and weights, you should be happy about this announcement. Incorporating as a nonprofit doesn't guarantee that an organization will act ethically, but it does make it more likely. Nonprofits have more re
72.
▲
by
vikp
4y ago
In this particular case (summarizing search results), it seems to work well. I think anchoring to known information helps avoid a lot of the issues (hallucinations, boring language, etc). You can check out an open source demo I made if you
73.
▲
by
vikp
4y ago
If you want to try this out today, I made an open source version using Google + GPT-3 - https://github.com/VikParuchuri/researcher . It works by getting Google results, finding the most relevant text chunks in the pages
74.
▲
by
vikp
4y ago
Hi HN - it's been getting hard for me to do research with Google. If I'm looking for the best smartphone, or the right Javascript framework, I have to wade through dozens of SEO-spam pages to find the answer. I've experimente
75.
▲
Show HN: Researcher – answer questions using Google and GPT-3
(github.com)
1 points
by
vikp
4y ago
|
1 comments
76.
▲
by
vikp
4y ago
I've had the exact same problem, and have experimented with Kagi, DDG, etc, to try to find better results. ChatGPT, in my opinion, is great for "how do I code X" type questions, but isn't so good at the types of queries
77.
▲
by
vikp
4y ago
It's doing abstractive summarization over the search results, using GPT-3. The pipeline is: - Search using Google - Run some filters to exclude SEO spam, etc. - Scrape the pages that are returned - Find chunks of text likely
78.
▲
by
vikp
4y ago
This is very cool! It's nice to see other implementations outside of ChatGPT. I noticed that it doesn't always cite sources, so I assume that some results are coming straight from an LLM, and some are web search + LLM. If you
79.
▲
by
vikp
4y ago
I don't think it's a huge lift to restrict a language model to "known" good facts from search results. And to have it cite sources. I made a proof of concept this weekend - https://github.com/VikParuchur
80.
▲
by
vikp
4y ago
I built a side project recently to help with this. It searches Google, then feeds the relevant results from the pages into GPT-3 to get a summary. It seems to be accurate so far in my testing - https://github.com/VikParuch
81.
▲
by
vikp
4y ago
Don't start a business unless it's a pain point you are personally invested in. Lifestyle-wise, it's a grind with mental health challenges along the way, as you pointed out. Salary-wise, it's unlikely that your own bus
82.
▲
by
vikp
4y ago
I haven't seen it specifically on HN, but SearX does what you described - https://searx.github.io/searx/ .
83.
▲
by
vikp
4y ago
Producing images of spectrograms is a genius idea. Great implementation! A couple of ideas that come to mind: - I wonder if you could separate the audio tracks of each instrument, generate separately, and then combine them. This could giv
84.
▲
by
vikp
4y ago
Large language models like GPT work by generating the probabilities for the next word in a sequence, given the previous words. You can make this purely deterministic (same sequence every time) by just selecting the word with the highest pro
85.
▲
by
vikp
4y ago
It's unclear to me how you could separate knowledge and reasoning: - Reasoning typically requires base knowledge to work from. A side effect of training reasoning is embedding knowledge into the model parameters. - Even if you offload
86.
▲
by
vikp
4y ago
I was curious to see an implementation, and I found this code for an earlier version of CFCs - https://github.com/raminmh/CfC .
87.
▲
by
vikp
7y ago
I haven't seen Withings Sleep mentioned - https://www.withings.com/us/en/sleep . This one is basically "set it and forget it". You put it under your mattress and calibrate it, then it tracks automa
88.
▲
by
vikp
8y ago
If you're only getting repaid when someone is employed and making over 50k, which seems to be the case from your website ("Upon completion of the program, students will pay 10% of their [pre-tax] salary for a five year period once
89.
▲
by
vikp
8y ago
This is very interesting, and I'm curious about the nature of the risk that you're taking on. It seems like risk would only become a factor for Lambda School if you're taking 10% of pre-tax income from graduates for the 5 chr
90.
▲
by
vikp
8y ago
We moved to Twist ( https://twistapp.com ) a few months ago after having similar issues with Slack. Twist is more forum-like, so you avoid the "I have to jump in now" feeling that a continuous chat stream gives you.
More ›