Ray Fernando
speaker
138 appearances
1 recordings
1 series
first heard Jan 2025
last heard Jan 2025
Ray Fernando’s voice in public audio — every appearance, attributed to the second.
Trend
recordings per month · last 12 monthsNo recordings in the last 12 months.Older appearances are listed below; set an alert to hear about the next one.
Appearances
Grok was able to get a distilled model. There's just so much demand. There's going to be even more demand for these chips. And yeah, this is just the beginning and I'm trying to figure out which provider can host this for me reliably so that I can do this for myself, but also share back and put this into apps for other people as well, because this is going to be This is great.
And I don't want the data to go to China. I just want the data to stay in North America. Or if I get a European container, I can do the European container and meet all their legal requirements that need to happen for that as well. So that's super exciting. Yeah.
Yeah, the pricing for fireworks, we can look this up real quick. I think it's about eight dollars per million tokens where normally I think ChatGPT was like 15 input and 60 dollars for output for O1 Pro. I can just double check that real quick. So pricing. Let's see.
Yeah. So 01 Pro or 01 API cost. Yeah. And the pricing.
That's exactly it. Yeah, they'll add up. Also, at the same time, OpenAI has currently promised that the 03 model will come out and the mini model will come out, which would be on par with this model. So that prices will also probably significantly drop as well because they just get more efficient with time. And so that'll be really interesting to see.
And, you know, I'm rooting for it because for me as a consumer, I want the power of all the intelligence. And to do these types of things, I think it's going to be pretty important. And I think to that note, I think it'd be interesting to show how if anyone hasn't found out about this. It's this thing on OpenAI's thing called the platform.openai.com.
So you just sign up for accounts or developer account. It's a little playground. And what you can do is actually hit this little generate star button. And so we can describe a prompt that we want for any type of model. And what it will do is reconfigure the prompt for the language model so that they can actually be more efficient at doing something.
So if there's a task that you do quite a bit, you're like, I just want, you know, please like make you know keywords for my Amazon listings and then hit like generate what it'll do is basically whoops sorry if we hit generate here and hit this at the very top right and hit update what it will do is actually reconfigure this prompt just from a one line type of thing to include more details so
we can take existing prompts and try to improve them through this mechanism. So as you can see, this is like how people get really nice long chains of thought or reasoning or outputs. So one way to think about these things is to first kind of put down like what instructions you want, what type of output do you really expect? maybe what you don't want as a good starting point.
And then that'll help you generate prompts that can be a little bit useful. A lot of the prompts that you're seeing that I spit out are basically things that have come out over time because of my use cases. It's like, okay, I want this instead of that. And so as an example, One of the things that I was thinking about was like, how do you verify like claims for a specific type of thing?
So if you have an article and how do you understand if something is actually true or not? So I have one here. This is information verification. And so one of the things that these models that I was currently showing you is that the Fireworks and the Grok are specific API endpoints. And right now there isn't like a specific web search thing that's currently tuned into them.
So if you want to do web search, you have to go through the deepseek.com route or the app. And keep in mind, you're also sending data into this container there. So sometimes you could just use it for public articles for things that you really don't care about. So if you're on DeepSeek, let's just see if they actually have stuff available here for us.
So you just go to the search thing, turn that on. I paste my prompt in and then what I'm going to go ahead and do is like maybe grab an article like the Techno Optimist Manifesto from Marc Andreessen. It's a very popular article and sometimes, you know, it's really long to read. There's like a lot of information here. And you're like, oh, my God, there's like so much in there.
It's like, how can I even get started with this thing? And how do I even verify the claims of this stuff? And I think this is probably sometimes like a good thing to start here. So what I do is I hit shift and enter and I put the article at the very top. And once I do that, I just go ahead and paste that in there and then just go ahead and hit send.
So what that's going to do for us is going to use the web search and try to look through, just like how we saw earlier that was happening with the API, every type of claim that's in there, try to see if they can search the internet for it and try to see if it can do anything. So this is going to try to do its thinking thing.
This is obviously very popular and DeepSeek, the website, is getting flooded with people because of basically it being free. And so, yeah, just keep in mind, like Greg says, like, yeah, I would not be putting taxes in there. I probably wouldn't be putting medical records in there.
Things that, you know, you don't want to see generated if somebody asks a question that's related to you, because that can be a little crazy. You'll be like, wow, all of a sudden my data is showing up somewhere. I was not expecting it. So Yeah, this is currently airing out. It's no surprise, and this is kind of why we were kind of thinking about doing some other alternatives here.
So yeah, I think that was set there. So I think another thing maybe that could be useful here is probably getting this stuff set up. So if you wanted to run this locally, maybe we can kind of go over that workflow. What do you think?
OK, OK, cool. Awesome. Let's do that. Yeah, I think that that'll be great. So in order, like for this section, in order for us to run this model locally, the best interface that I found, bar none, is something called Open Web UI. So Open Web UI, it's really quick to get started. All you need is to download Docker. So just go to docker.com. And you can just download the desktop app.
So download the one that you need for your machine. If you're running an Apple M machine, which I am doing, I just download the Apple Silicon version and accordingly. So once you get that installed, it's going to present to you a user interface like a dashboard here. And so that's kind of something that you'll have kind of going.
Showing 41–60 of 138 · page 3 of 7
← Previous
Next →