Astra will initiate test suites, find one more thing independently while its running, reinitiate complete test suite after fixing it, then find one more thing, then test again. Easy to burn through GH actions minutes if…
I wonder too if in training for long horizon tasks agents become worse team players, good at orchestrating subagents they are trained to use, but worse as an agent within an external multi-agent orchestration system or…
In SWE I've found gpt-6-astra (high) inconsistent and oddly focused on overtly taking responsibility for mistakes it made rather than prioritizing concrete steps to rectify problems. Such steps once elicited are often…
I also temporarily had agy as a harness in my orchestrated workflow until I learned about the account bans. IANAL but frontier models examining Google's TOS suggested my use could be in violation. Haven't touched agy…
I've tested Omnigent superficially, attracted to its thinking around policy, governance, sandboxing, and ui. But it's still alpha at present. I forked its Polly model and got working a somewhat more complex multiagent…
I can also vouch for magnesium and the l-threonate variant. I take both before bed along with glycine powder, phosphatidyl-serine, l-theanine, l-tryptophan, ashwagandha, and saffron. No melatonin, no sleeping pills.…
Astra will initiate test suites, find one more thing independently while its running, reinitiate complete test suite after fixing it, then find one more thing, then test again. Easy to burn through GH actions minutes if…
I wonder too if in training for long horizon tasks agents become worse team players, good at orchestrating subagents they are trained to use, but worse as an agent within an external multi-agent orchestration system or…
In SWE I've found gpt-6-astra (high) inconsistent and oddly focused on overtly taking responsibility for mistakes it made rather than prioritizing concrete steps to rectify problems. Such steps once elicited are often…
I also temporarily had agy as a harness in my orchestrated workflow until I learned about the account bans. IANAL but frontier models examining Google's TOS suggested my use could be in violation. Haven't touched agy…
I've tested Omnigent superficially, attracted to its thinking around policy, governance, sandboxing, and ui. But it's still alpha at present. I forked its Polly model and got working a somewhat more complex multiagent…
I can also vouch for magnesium and the l-threonate variant. I take both before bed along with glycine powder, phosphatidyl-serine, l-theanine, l-tryptophan, ashwagandha, and saffron. No melatonin, no sleeping pills.…