Open invitation: I'll help you think through this multimodal RAG interview, if you let me pitch you @pixeltablehq so you don't have to (re)build all this plumbing from scratch when you get the job 🍞
pixeltable.com/toast
interview question I like - multi-modal RAG w/ an agent
can go high level & also rabbit hole into context engineering, tokenization, harness design, search, how models are trained, etc
Setup:
there’s a folder with multi-modal data - markdown files, PDFs, image files, video
In case you missed it, this is a great prompt for writing a good ticket whether or not you use @linear. I do this during every video tutorial I create for the development team @pixeltablehq.
Not sure how it exactly happens in your case.
1. we use linear agent a lot to create issues from slack/elsewhere so it's already tuned to do in less verbose way
2. There is the option to run loops to new issues to somehow refactor them to me more human readable
3. Internally
This is the best reason to use @pixeltablehq - our customers start with their data model and a Pixeltable schema that includes multimodal types. Then you can start hurling your tokens into your tables.
building prototypes of mockups, schemas, data models, proof of concepts, etc. is the best way to avoid spending tons of tokens before realizing you don't want the output
Today, we’re launching Reve 2.0, the best 4K image model in the world.
We invented a new way to generate and edit any image using precise layouts. For the first time, it’s possible to create images you can touch.