Technology

64937 readers

4587 users here now

This is a most excellent place for technology news and articles.

Our Rules

Follow the lemmy.world rules.
Only tech related content.
Be excellent to each other!
Mod approved content bots can post up to 10 articles per day.
Threads asking for personal tech support may be deleted.
Politics threads may be removed.
No memes allowed as posts, OK to post as comments.
Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
Check for duplicates before posting, duplicates may be removed
Accounts 7 days and younger will have their posts automatically removed.

Approved Bots

founded 2 years ago

MODERATORS

[email protected]

313

Meta says you can’t turn off its new AI tool on Facebook, Instagram (globalnews.ca)

submitted 10 months ago by [email protected] to c/[email protected]

43 comments fedilink hide all child comments

you are viewing a single comment's thread
view the rest of the comments

[–] [email protected] 9 points 10 months ago (2 children)

At least I can run Llama 3 entirely locally.

[–] [email protected] 4 points 10 months ago

I just discovered how easy ollama and open webui are to set up so I've been using llama3 locally too, it was like 20 lines in docker compose, and although I've been using gpt3.5 on and off for a long time I'm much more comfortable using models run locally so I've been playing with it a lot more. It's also cool being able to easily switch models at any point during a conversation. I have like 15 models downloaded, mostly 7b and a few 13b models and they all run fast enough on CPU and generate slightly slower than reading speed and only take ~15-30 seconds to start spitting out a response.

Next I want to set up a vscode plugin so I can use my own locally run codegen models from within vscode.

[–] [email protected] 3 points 10 months ago (1 children)

I tried llamas when they were initially released, and it seems like training took garbage amounts of GPU. Did that change?

[–] [email protected] 2 points 10 months ago

Look into quantised models (like gguf format) these significantly reduce the amout of memory needed and speed up computation time at the expense of some quality. If you have 16GB of rm or more you can run decent models locally without any gpu, though your speed will be more like 1 word a second than chatgpt speeds