Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
fxwin
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
fxwin
2mo ago
Yea it's a shitty environment, it's just a "oh sweet summer child" kind of thing to say :) I agree that X forwarding is a terrible way for GUI apps, but the OP read like they were just looking at logs or plots/visua
32.
▲
by
fxwin
2mo ago
OP is not ready for corporate environments if "VPN + ssh + 2FA" already "feels like too much" lol I also wonder how they didn't come across X Forwarding at all? When i used a remote machine for my thesis that was on
33.
▲
by
fxwin
2mo ago
Nobody is saying that they use "so little water and power", just that the average person has no clue about the orders of magnitude at which anything outside of a private household uses these things, including things they previousl
34.
▲
by
fxwin
2mo ago
> Earlier this week, we detected and responded to an intrusion into part of our production infrastructure. This one was different from anything we had handled before in one important way: it was driven, end to end, by an autonomous AI ag
35.
▲
by
fxwin
2mo ago
im not sure which part of my summary you take issue with then? the reason for study 1b is not relevant to answer the question asked in the comment i replied to
36.
▲
by
fxwin
2mo ago
You can read the study here: https://osf.io/preprints/psyarxiv/5y6m4_v1 TLDR: They studied both cases (Access to a LLM/Chat interface which gives a wrong answer when asked, and access to pregenerated (wrong)
37.
▲
by
fxwin
2mo ago
definitely the latter, it is even referenced in the foreword: > Its goal is not to be exhaustive, but rather minimalist and easy to read. For this reason, it follows the format of The Little Book of Deep Learning [Fleuret 2023]. Its tone
38.
▲
by
fxwin
2mo ago
looks like they have since fixed the dataset, hope they release an updated report (either without the leaked evals or a new checkpoint/weights) https://huggingface.co/datasets/AIML-TUDA/QA-base/commit
39.
▲
by
fxwin
2mo ago
I think the real challenge is to figure out wtf you're actually supposed to do from all this flowery over the top claudespeak
40.
▲
by
fxwin
2mo ago
> The reality is people don't always care if a human poured their heart and soul into something. That's fine, but I don't think the author would suggest writing e.g. library documentation by hand. It's clearly advice
41.
▲
by
fxwin
2mo ago
That was a fun read! I caught myself almost skimming the first part until i got to the mirrored paragraph, and slowed down significantly after that to read more deliberately. I'm not sure how much actual advice one can take from this e
42.
▲
by
fxwin
2mo ago
>For four months, no frontier model beat Claude Opus in our production evals What do you think that is referring to?
43.
▲
by
fxwin
2mo ago
> The Pi 3B+ seems just enough to render some simple animations. > I think a Pi 4 might be a good sweet spot between processing power and price, I know this isn't exactly a serious product and more of a gadget/gimmick but m
44.
▲
by
fxwin
3mo ago
I would expect they have production based datasets they evaluate new models against.
45.
▲
by
fxwin
3mo ago
Tried exploring a small project i built with CC, but i don't see anything in the tree/terrain view (The edits/reads/writes do show up in the timeline). The project itself doesn't exist on my drive anymore, is that a
46.
▲
by
fxwin
3mo ago
fwiw o1, o3 and 4.1 also give the correct answer (without web search)
47.
▲
by
fxwin
3mo ago
I'm not sure which chatbots you used, but OpenAI's o3, o1 and 4.1 get it right first try (used in the Playground without web search or any other tools).
48.
▲
by
fxwin
3mo ago
Sure, but that's the nature of language (which is also why i put "understand" in quotation marks. I usually follow it up with "whatever that means" lol) . I think in this case, it carries with it implicit properties
49.
▲
by
fxwin
3mo ago
I think "(intelligent) language understander" is an apt term. It contains within it the fact that these models are mainly trained on text, and "understand" it beyond a simple token-by-token level (i.e. their latent space
50.
▲
by
fxwin
3mo ago
Have they retracted it? My understanding was simply that they released results with more recent data, not that this study itself was flawed (and their website doesn't mention a retraction either)
51.
▲
by
fxwin
3mo ago
And i would consider this as an agent doing things under human instruction
52.
▲
by
fxwin
3mo ago
I'm not sure which part of my comment comes across as fence sitting to you, but to clarify: - I think there are good and useful ways to use AI in art creation - I also think that AI (especially end to end creation of full songs) is a m
53.
▲
by
fxwin
3mo ago
I work with coding agents every day, I don't think they have ever started working on a project without me telling them to
54.
▲
by
fxwin
3mo ago
> Laws and rules exist to serve humans, not machines. Machines don't go out on their own to create and upload music, they do so under human instruction, so their output should be policed the same way we police other machine generate
55.
▲
by
fxwin
3mo ago
The issue i have with it will depend heavily on implementation, i can see cases where songs that i would consider "produced and written" by people don't qualify for royalties under Tidal's guidelines. (I intentionally le
56.
▲
by
fxwin
3mo ago
> Tidal will accept AI-generated music. > Tidal will hold AI-generated music to a higher standard of content integrity. We will not tolerate AI-generated music that exploits an individual’s or group’s music, name or likeness, deceives
57.
▲
by
fxwin
3mo ago
fwiw aleph alpha have been around since 2019
58.
▲
by
fxwin
3mo ago
> Those are specs that belong on a laptop or a lowend console offering like the Series S. Unfortunately, valve (and we consumers) have to recalibrate our understanding of which prices qualify as "insulting".
59.
▲
by
fxwin
3mo ago
Me, a german, looking at my CAX11: this doesn't seem too bad Americans:
60.
▲
by
fxwin
4mo ago
I think it helps his credibility that he has been working with and speaking positively about AI assisted mathematics (especially for formalizing proofs) for over a year now . I'm sure he isn't unbiased, but as far as spokespeople
More ›