Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
gregschoeninger
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
gregschoeninger
2mo ago
We've been working for a few years on a VCS that is starting to get adopted by creative studios - called "Oxen" It's open source: https://github.com/oxen-AI/oxen Most users are building their own in
2.
▲
by
gregschoeninger
4mo ago
We're also working on an open source large asset versioning tool called "oxen" - https://github.com/Oxen-AI/Oxen Would love any feedback on it or contributions if people are interested :)
3.
▲
by
gregschoeninger
6mo ago
We're working on this project to help with the non-text file and large file problem: https://github.com/Oxen-AI/Oxen Started with the machine learning use case for datasets and model weights but seeing a lot of tr
4.
▲
by
gregschoeninger
2y ago
Over the past ~1.5 years I've been running a research paper club where we dive into interesting/foundational papers in AI/ML. So we naturally have come across a lot of the papers that lead up to DeepSeek-R1. While diving into
5.
▲
by
gregschoeninger
2y ago
Hey all, If you haven't seen the Oxen project yet, we have been building an open source unstructured data version control tool. We were inspired by the idea of making large machine learning datasets living & breathing assets that p
6.
▲
by
gregschoeninger
2y ago
Maintainer of Oxen here, we initially built Oxen because DVC was pretty painfully slow to work with, and had a lot of extra bells and whistles that we didn’t need. Under the hood we optimized the merkle tree structure, hashing algorithms, n
7.
▲
Paper Club: How Flux.1 models work under the hood
(oxen.ai)
2 points
by
gregschoeninger
2y ago
|
1 comments
8.
▲
by
gregschoeninger
2y ago
Hey all, With Black Forest Labs’ Flux.1 variants being the current state of the art for image gen, we’re doing a technical dive into a few paper that inspired the work, starting with: Scaling Rectified Flow Transformers for High-Resolution
9.
▲
Using Llama3.1 405B to generate political synthetic data
(oxen.ai)
5 points
by
gregschoeninger
2y ago
|
3 comments
10.
▲
by
gregschoeninger
2y ago
We thought it'd be interesting to see what political biases Llama 3.1 405B has by generating a bunch of "spam" or "ham" messages with it. We started with 5 hand crafted messages and let the LLM take it from there en
11.
▲
Fine Tuning a Diffusion Transformer (DiT) from a Single YouTube Video
(oxen.ai)
4 points
by
gregschoeninger
2y ago
|
2 comments
12.
▲
by
gregschoeninger
2y ago
Hey all, We were messing around with PixArt as a way to fine tune DiT's for image generation. I was pretty impressed with the results and thought I'd share. https://www.oxen.ai/ox/PixArtTutorial In this examp
13.
▲
How to train diffusion for text from scratch
(ghost.oxen.ai)
1 points
by
gregschoeninger
2y ago
|
1 comments
14.
▲
by
gregschoeninger
2y ago
Hey all, I thought the paper “Discrete Diffusion Modeling by Estimating the ratios of the Data Distribution” was a pretty cool idea, so decided to dive deep into the code, strip it down so I could understand it, then train some models from
15.
▲
Instruct-Tuning BitNet 1.58
(github.com)
4 points
by
gregschoeninger
2y ago
|
2 comments
16.
▲
by
gregschoeninger
2y ago
This is work done for our arxiv dive paper club where we dive into research papers and implement code to see how the models work in practice. We have some internal use cases for BitNets so thought we'd share the work as we go along. En
17.
▲
by
gregschoeninger
3y ago
We used an A10 with 24GB of VRAM, this was enough for PEFT on Mistral-7B
18.
▲
by
gregschoeninger
3y ago
The goal is to iteratively create training data and add it to its own training set. The LLM acts as its own judge and scores its own responses to decide if it should add the data. It’s expensive to have a human in the loop labeling preferen
19.
▲
Show HN: Implementation of the "Self-Rewarding Language Models" Paper by MetaAI
(github.com)
23 points
by
gregschoeninger
3y ago
|
5 comments
20.
▲
by
gregschoeninger
3y ago
Hey all, After reading the Self-Rewarding Language Models paper by the team at Meta, it felt very approachable and reproducible, so we spent some time implementing it. The scripts provided take any base model and put it in a loop of: 1) Sup
21.
▲
"Road to Sora" Paper Reading List
(oxen.ai)
32 points
by
gregschoeninger
3y ago
|
1 comments
22.
▲
by
gregschoeninger
3y ago
Hey all, Have been diving into the Sora technical report for our paper club on Friday, and decided it would be nice to have a reading list of the background papers need to fully grok everything that is going on in that technical report - ea
23.
▲
Show HN: Oxen.ai – Data Diff tool to quickly find changes in CSV, parquet, etc.
(docs.oxen.ai)
4 points
by
gregschoeninger
3y ago
|
0 comments
24.
▲
Guide to the Mamba architecture that claims to be a replacement for Transformers
(blog.oxen.ai)
5 points
by
gregschoeninger
3y ago
|
2 comments
25.
▲
by
gregschoeninger
3y ago
Been diving deep into the Mamba paper and put together my notes here: https://blog.oxen.ai/mamba-linear-time-sequence-modeling-wit... Took me awhile to wrap my head around some of the terminology, so hopefully this helps an