Isn't ChatGPT benchmaxxing, then ? Responding "hmm…" isn't actually responding and latency should time to first relevant phoneme.
I have zero interest in world knowledge for my LLMs but this got me wondering : are there RAGs for that kind of data ? How could a LLM like DeepSeek-v4-flash-vision-exp accurately answer you question with an indexed…
Indeed ! LLM are creative writers, not journalists. Relying on overfitting for factual accuracy in not tenable. I don't understand why grounded RAG with judges in not the norm.
You understand that drones are used for precision strikes instead of indiscriminate bombing, right ? It's the opposite of "sponsoring a genocide" (if there actually ever was one happening…)
Everything is food and air. There is no such thing a "human labor" without food and air to sustain it. I'm not sure what your point is.
BTW, I don't understand why having the driver converting AC to DC inside the LED spots is the default. For a new house, it seems to make sense to me to have at least one external driver for a ceiling of spots, if not…
I presume this works (will work) also for JupyterLite that is based on Pyodide ? Would be great if it helped getting the latest OpenCV-python version [0] and it's dnn goodies being available on a zero-install client…
OpenCV being in the list of Pyodide modules [0] was the biggest boon for my online teaching experience because remotely dealing with install woes (corporate proxies & cie) was a show stopper for regular Python. I'm…
Nice ! My most pressing request for VSS would be efficient binary vectors : is this on the table ?
Only those who don't care/know about prompt processing speed are buying Macs for LLM inference.
For LLM inference, I don't think the PCIe bandwidth matters much and a GPU could improve greatly the prompt processing speed.
Indeed, recent Flash Attention is a pain point for non CUDA.
The idea is presumably that you would "sell" at an artificially low price.
> The Soviets had the […]first woman,[…] That is quite the claim !
You forgot the "/s", or do you actually believe that it's capitalism's fault is a mother taking care of her children is "unpaid labor" ?
"is hard" ≠ "sucks"
Most interesting ! Would you mind sharing the prompt and the resulting CLAUDE.md file ? Thx !
IMO, it would be more interesting to have a 3-way comparison of price/performance between DeepSeek 671b running on : 1. M3 Ultra 512 2. AMD Epyc (which Gen ? AVX512 and DDR5 might make a difference in both performance…
DeepSeek is not a model.Which model did you use (v3 ? R1 ? a distillation ?) at which quantization ?
Nice ! Is it possible to connect to an in browser DB like WASM DuckDB https://duckdb.org/docs/api/wasm/overview.html or https://github.com/babycommando/entity-db ? That would be most useful imho !
It seems that this aims to refutes claims for inaction with facts about spending money. However, the high speed rail project or homelessness management seem to show that in California, $$$ spent doesn't always imply…
How do we know that this extension can be trusted ?
«Unfortunately, I have only seen 3 models, 3B or over, handle RAG.» I would love to know which are these 3 models, especially if they can perform grounded RAG. If you have models (and their grounded RAG prompt formats)…
Just played a bit with it. Were you working with ASCII ? This example didn't work for you ? https://github.com/jfalcou/eve/blob/a141ba93048bb2916c2157a9...
Counterpoint : your message is not synthetic data and will contribute to lots of LLMs saying the same. Many such cases ? (It seems to me obvious that a fgrep would sanitize synthetic data obtained from competitors.)
Isn't ChatGPT benchmaxxing, then ? Responding "hmm…" isn't actually responding and latency should time to first relevant phoneme.
I have zero interest in world knowledge for my LLMs but this got me wondering : are there RAGs for that kind of data ? How could a LLM like DeepSeek-v4-flash-vision-exp accurately answer you question with an indexed…
Indeed ! LLM are creative writers, not journalists. Relying on overfitting for factual accuracy in not tenable. I don't understand why grounded RAG with judges in not the norm.
You understand that drones are used for precision strikes instead of indiscriminate bombing, right ? It's the opposite of "sponsoring a genocide" (if there actually ever was one happening…)
Everything is food and air. There is no such thing a "human labor" without food and air to sustain it. I'm not sure what your point is.
BTW, I don't understand why having the driver converting AC to DC inside the LED spots is the default. For a new house, it seems to make sense to me to have at least one external driver for a ceiling of spots, if not…
I presume this works (will work) also for JupyterLite that is based on Pyodide ? Would be great if it helped getting the latest OpenCV-python version [0] and it's dnn goodies being available on a zero-install client…
OpenCV being in the list of Pyodide modules [0] was the biggest boon for my online teaching experience because remotely dealing with install woes (corporate proxies & cie) was a show stopper for regular Python. I'm…
Nice ! My most pressing request for VSS would be efficient binary vectors : is this on the table ?
Only those who don't care/know about prompt processing speed are buying Macs for LLM inference.
For LLM inference, I don't think the PCIe bandwidth matters much and a GPU could improve greatly the prompt processing speed.
Indeed, recent Flash Attention is a pain point for non CUDA.
The idea is presumably that you would "sell" at an artificially low price.
> The Soviets had the […]first woman,[…] That is quite the claim !
You forgot the "/s", or do you actually believe that it's capitalism's fault is a mother taking care of her children is "unpaid labor" ?
"is hard" ≠ "sucks"
Most interesting ! Would you mind sharing the prompt and the resulting CLAUDE.md file ? Thx !
IMO, it would be more interesting to have a 3-way comparison of price/performance between DeepSeek 671b running on : 1. M3 Ultra 512 2. AMD Epyc (which Gen ? AVX512 and DDR5 might make a difference in both performance…
DeepSeek is not a model.Which model did you use (v3 ? R1 ? a distillation ?) at which quantization ?
Nice ! Is it possible to connect to an in browser DB like WASM DuckDB https://duckdb.org/docs/api/wasm/overview.html or https://github.com/babycommando/entity-db ? That would be most useful imho !
It seems that this aims to refutes claims for inaction with facts about spending money. However, the high speed rail project or homelessness management seem to show that in California, $$$ spent doesn't always imply…
How do we know that this extension can be trusted ?
«Unfortunately, I have only seen 3 models, 3B or over, handle RAG.» I would love to know which are these 3 models, especially if they can perform grounded RAG. If you have models (and their grounded RAG prompt formats)…
Just played a bit with it. Were you working with ASCII ? This example didn't work for you ? https://github.com/jfalcou/eve/blob/a141ba93048bb2916c2157a9...
Counterpoint : your message is not synthetic data and will contribute to lots of LLMs saying the same. Many such cases ? (It seems to me obvious that a fgrep would sanitize synthetic data obtained from competitors.)