Another hypothesis is that OpenAI's internal web sandbox (WebCache, also used during HF attack [1]) was too narrow/restrictive for the agents' purposes. During collusion-wiki, it was hypothesized that the sandbox would…
You need to separate memory capacity and bandwidth. Looping decreases memory capacity/FLOP but not bytes loaded/FLOP, since weights need to be loaded again for the 2nd pass. Plus (depending on the method used) capacity…
What LLM-specific hardware improvements should one expect? Seems to me that LLM inference is simple architecturally (matmul et al) so most scaling in hardware should come from general improvements (memory BW, packaging,…
Consider that a lot of the resistance to just existing is external expectations. I feel like I should always be doing something because I'm pressured to, or that small things bother me because I feel I need to perform.…
Another hypothesis is that OpenAI's internal web sandbox (WebCache, also used during HF attack [1]) was too narrow/restrictive for the agents' purposes. During collusion-wiki, it was hypothesized that the sandbox would…
You need to separate memory capacity and bandwidth. Looping decreases memory capacity/FLOP but not bytes loaded/FLOP, since weights need to be loaded again for the 2nd pass. Plus (depending on the method used) capacity…
What LLM-specific hardware improvements should one expect? Seems to me that LLM inference is simple architecturally (matmul et al) so most scaling in hardware should come from general improvements (memory BW, packaging,…
Consider that a lot of the resistance to just existing is external expectations. I feel like I should always be doing something because I'm pressured to, or that small things bother me because I feel I need to perform.…