7 ms·
They've got a console for it as well, https://www.meta.ai/ https://www.meta.ai/ And announcing a lot of integration across the Meta product suite, https://abou
by bbig 2y ago
They've got a console for it as well,
https://www.meta.ai/ https://www.meta.ai/
And announcing a lot of integration across the Meta product suite,
https://about.fb.com/news/2024/04/meta-ai-assistant-built-with-llama-3/ https://about.fb.com/news/2024/04/meta-ai-assistant-built-wi...
Neglected to include comparisons against GPT-4-Turbo or Claude Opus, so I guess it's far from being a frontier model. We'll see how it fares in the LLM Arena.
- krackers 2y agoAre there an stats on if llama 3 beats out chatgpt 3.5 (the free one you can use)?
- throwup238 2y ago> And announcing a lot of integration across the Meta product suite, ... That's ominous...
- iosjunkie 2y agoSpending millions/billions to train these models is for a reason and it's not just for funsies.
- nickthegreek 2y agoAnd they even allow you to use it without logging in. Didnt expect that from Meta.
- visarga 2y agoDoesn't work for me, I'm in EU.
- mvkel 2y agoProbably bc they're violating gdpr
- salil999 2y agoI do see on the bottom left: Log in to save your conversation history, sync with Messenger, generate images and more.
- zitterbewegung 2y agoThink they meant it can be used without login.
- applecrazy 2y agoI imagine that is to compete with ChatGPT, which began doing the same.
- lairv 2y agoNot in the EU though
- sega_sai 2y agoor the UK
- HarHarVeryFunny 2y agoYeah, but not for image generation unfortunately I've never had a FaceBook account, and really don't trust them regarding privacy
- zingelshuher 2y agohad to upvote this
- sdesol 2y agoI had the same reaction, but when I saw the thumbs up and down icon, I realized this was a smart way to crowd source validation data.
- unshavedyak 2y agoWhich indicates that they get enough value out of logged ~in~ out users. Potentially they can identify you without logging in, no need to. But also ofc they get a lot of value by giving them data via interacting with the model.
- hakdbha 2y ago[dead]
- MichaelCharles 2y agoBut not from Japan, and I assume most other non-English speaking countries.
- mvkel 2y ago1. Free rlhf 2. They cookie the hell out of you to breadcrumb your journey around the web. They don't need you to login to get what they need, much like Google
- eggdaft 2y agoDo they really need “free RLHF”? As I understand it, RLHF needs relatively little data to work and its quality matters - I would expect paid and trained labellers to do a much better job than Joey Keyboard clicking past a “which helped you more” prompt whilst trying to generate an email.
- spi 2y agoVariety matters a lot. If you pay 1000 trained labellers, you get 1000 POVs for a good amount of money, and likely can't even think of 1000 good questions to have them ask. If you let 1000000 people give you feedback on random topics for free, and then pay 100 trained people to go through all of that and only retain the most useful 1%, you get much ten times more variety for a tenth of the cost. Of course numbers are pretty random, but it's just to give an idea of how these things scale. This is my experience from my company's own internal -deep learning but not LLM- models to train which we had to buy data instead of collecting it. If you can't tap into data "from the wild" -in our case, for legal reason- you can still get enough data (if measured in GB), but it's depressingly more repetitive, and that's not quite the same thing when you want to generalize.
- mvkel 2y agoAbsolutely. Modern captchas are self driving object labelers; you just need a few to "agree" to know what the right answer is.
- dizhn 2y agoWe should agree on a different answer for crosswalk and traffic light and mess it up for them.
- yakorevivan 2y ago
- CuriouslyC 2y agoThey didn't compare against the best models because they were trying to do "in class" comparisons, and the 70B model is in the same class as Sonnet (which they do compare against) and GPT3.5 (which is much worse than sonnet). If they're beating sonnet that means they're going to be within stabbing distance of opus and gpt4 for most tasks, with the only major difference probably arising in extremely difficult reasoning benchmarks. Since llama is open source, we're going to see fine tunes and LoRAs though, unlike opus.
- danielhanchen 2y agoOn the topic of LoRAs and finetuning, have a Colab for LoRA finetuning Llama-3 8B :) https://colab.research.google.com/drive/135ced7oHytdxu3N2DNe1Z0kqjyYIkDXp?usp=sharing https://colab.research.google.com/drive/135ced7oHytdxu3N2DNe...
- htrp 2y agoML Twitter was saying that they're working on a 400B parameter version?
- mkl 2y agoMeta themselves are saying that: https://ai.meta.com/blog/meta-llama-3/ https://ai.meta.com/blog/meta-llama-3/
- blackeyeblitzar 2y agoLlama is open weight, not open source. They don’t release all the things you need to reproduce their weights.
- mananaysiempre 2y agoNot really that either, if we assume that “open weight” means something similar to the standard meaning of “open source”—section 2 of the license discriminates against some users, and the entirety of the AUP against some uses, in contravention of FSD #0 (“The freedom to run the program as you wish, for any purpose”) as well as DFSG #5&6 = OSD #5&6 (“No Discrimination Against Persons or Groups” and “... Fields of Endeavor”, the text under those titles is identical in both cases). Section 7 of the license is a choice of jurisdiction, which (in addition to being void in many places) I believe was considered to be against or at least skirting the DFSG in other licenses. At best it’s weight-available and redistributable.
- resource_waste 2y ago[flagged]
- SV_BubbleTime 2y agoSorry, still too sexy. Can’t have that.
- deleted 2y ago[deleted]
- SOVIETIC-BOSS88 2y agoWe are living in a post Dan Schneider world. Feet are off the table.
- resource_waste 2y agoI think nsfw stats bursted that bubble, not danny.
- sebastiennight 2y agoWell thanks then. Some of us eat on this table you know
- visarga 2y agoGPT-3.5 rejected to extract data from a German receipt because it contained "Women's Sportswear", sent back a "medium" severity sexual content rating. That was an API call, which should be less restrictive.
- freedomben 2y agoI haven't tried Llama 3 yet, but Llama 2 is indeed extremely "safe." (I'm old enough to remember when AI safety was about not having AI take over the world and kill all humans, not when it might offend a Puritan's sexual sensibilities or hurt somebody's feelings, so I hate using the word "safe" for it, but I can't think of a better word that others would understand). It's not quite as bad as Gemini, but in the same class where it's almost not useful because so often it refuses to do anything except lecture. Still very grateful for it, but I suspect the most useful model hasn't happened yet.
- deleted 2y ago[deleted]
- schleck8 2y ago> Neglected to include comparisons against GPT-4-Turbo or Claude Opus, so I guess it's far from being a frontier model Yeah, almost like comparing a 70b model with a 1.8 trillion parameter model doesn't make any sense when you have a 400b model pending release.
- cjbprime 2y ago(You can't compare parameter count with a mixture of experts model, which is what the 1.8T rumor says that GPT-4 is.)
- matsemann 2y ago> Meta AI isn't available yet in your country Where is it available? I got this in Norway.
- niek_pas 2y agoGot the same in the Netherlands.
- flemhans 2y agoProbably the EU laws are getting too draconian. I'm starting to see it a lot.
- sa-code 2y agoEU actually has the opposite of draconian privacy laws. It's more that meta doesn't have a business model if they don't intrude on your privacy
- mrtranscendence 2y agoWell, exactly, and that's why IMO they'll end up pulling out the EU. There's barely any money in non-targeted ads.
- sebastiennight 2y agoIf by "barely any money", you mean "all the businesses in the EU will still give you all their money as long as you've got eyeballs", then yes.
- ben_w 2y agoFacebook has shown me ads for both dick pills and breast surgery, for hyper-local events in town in a country I don't live in, and for a lawyer who specialises in renouncing a citizenship I don't have. At this point, I think paying Facebook to advertise is a waste of money — the actual spam in my junk email folder is better targeted.
- josh-sematic 2y agoThey also stated that they are still training larger variants that will be more competitive: > Our largest models are over 400B parameters and, while these models are still training, our team is excited about how they’re trending. Over the coming months, we’ll release multiple models with new capabilities including multimodality, the ability to converse in multiple languages, a much longer context window, and stronger overall capabilities.
- glenstein 2y agoAnyone have any informed guesstimations as to where we might expect a 400b parameter model for llama 3 to land benchmark wise and performance wise, relative to this current llama 3 and relative to GPT-4? I understand that parameters mean different things for different models, and llama two had 70 b parameters, so I'm wondering if anyone can contribute some guesstimation as to what might be expected with the larger model that they are teasing?
- ZiiS 2y agoThey are aiming to beat the current GPT4 and stand a fair chance, they are unlikly to hold the crown for long.
- glenstein 2y agoRight because the very little I've heard out of Sam Altman this year hinting at future updates suggests that there's something coming before we turn our calendars to 2025. So equaling or mildly exceeding GPT-4 will certainly be welcome, but could amount to a temporary stint as king of the mountain.
- llm_trw 2y agoThis is always the case. But the fact that open models are beating state of the art from 6 months ago is really telling just how little moat there is around AI.
- jamesgpearce 2y agoThat realtime `/imagine` prompt seems pretty great.
- geepytee 2y agoAlso added Llama 3 70B to our coding copilot https://www.double.bot https://www.double.bot if anyone wants to try it for coding within their IDE and not just chat in the console
- 8n4vidtmkvmk 2y agoCan we stop referring to VS Code as "their IDE"? Do you support any other editors? If the list is small, just name them. Not everyone uses or likes VS Code.
- DresdenNick 2y agoDone. Anything else?
- erhaetherth 2y agoNo, actually. Thank you for that. Your "Double vs. Github Copilot" page is great. I've signed up for the Jetbrains waitlist.
- rdez6173 2y agoDouble seems more like a feature than a product. I feel like Copilot could easily implement those value-adds and obsolete this product. I also don't understand why I can't bring my own API tokens. I have API keys for OpenAI, Anthropic, and even local LLMs. I guess the "secret" is in the prompting that is being done on the user's behalf. I appreciate the work that went into this, I just think it's not for me.
- doakes 2y agoThat was fast! I've really been enjoying Double, thanks for your work.
- ionwake 2y agoCool thanks! Will try
- dawnerd 2y agoTried a few queries and was surprised how fast it responded vs how slow chatgpt can be. Responses seemed just as good too.
- gliched_robot 2y agoInference speed is not a great metric given the horizontal scalability of LLMs.
- jaimex2 2y agoBecause no one is using it
- dazuaz 2y agoI'm based on LLaMA 2, which is a type of transformer language model developed by Meta AI. LLaMA 2 is a more advanced version of the original LLaMA model, with improved performance and capabilities. I'm a specific instance of LLaMA 2, trained on a massive dataset of text from the internet, books, and other sources, and fine-tuned for conversational AI applications. My knowledge cutoff is December 2022, and I'm constantly learning and improving with new updates and fine-tuning.
- davidmurdoch 2y agoAre you trying to say you are a bot?
- Aaron2222 2y agoThat's the response they got when asking the https://www.meta.ai/ https://www.meta.ai/ web console what version of LLaMA it is.
- salesynerd 2y agoStrange. The Llama 3 model card mentions that the knowledge cutoff dates are March 2023 for the 8B version and December 2023 for the 70B version (https://github.com/meta-llama/llama3/blob/main/MODEL_CARD.md https://github.com/meta-llama/llama3/blob/main/MODEL_CARD.md)
- gliched_robot 2y agoMaybe a typo?
- glenstein 2y agoI suppose it could be hallucinations about itself. I suppose it's perfectly fair for large language models not necessarily to know these things, but as far as manual fine tuning, I think it would be reasonable to build models that are capable of answering questions about which model they are, their training date, their number of training parameters, and how they are different from other models, etc. Seems like it would be helpful for it to know and not have to try to do its best guess and potentially hallucinate. Although in my experience Llama 3 seemed to know what it was, but generally speaking it seems like this is not necessarily always the case.
- LrnByTeach 2y agoLosers & Winners from Llama-3-400B Matching 'Claude 3 Opus' etc.. Losers: - Nvidia Stock : lid on GPU growth in the coming year or two as "Nation states" use Llama-3/Llama-4 instead spending $$$ on GPU for own models, same goes with big corporations. - OpenAI & Sam: hard to raise speculated $100 Billion, Given GPT-4/GPT-5 advances are visible now. - Google : diminished AI superiority posture Winners: - AMD, intel: these companies can focus on Chips for AI Inference instead of falling behind Nvidia Training Superior GPUs - Universities & rest of the world : can work on top of Llama-3
- gliched_robot 2y agoDisagree on Nvidia, most folks fine-tune model. Proof: there are about 20k models in huggingface derived from llama 2, all of them trained on Nvidia GPUs.
- vineyardmike 2y agoI also disagree on Google... Google's business is largely not predicated on AI the way everyone else is. Sure they hope it's a driver of growth, but if the entire LLM industry disappeared, they'd be fine. Google doesn't need AI "Superiority", they need "good enough" to prevent the masses from product switching. If the entire world is saturated in AI, then it no longer becomes a differentiator to drive switching. And maybe the arms race will die down, and they can save on costs trying to out-gun everyone else.
- 2y ago
- niutech 2y agoWhy does Meta embed a 3.5MB animated GIF (https://about.fb.com/wp-content/uploads/2024/04/Meta-AI-Expasion_Header.gif https://about.fb.com/wp-content/uploads/2024/04/Meta-AI-Expa...) on their announcement post instead of much smaller animated WebP/APNG/MP4 file? They should care about users with low bandwidth and limited data plan.