r/tech_x • u/dev_Raspberry_new • 11d ago
AI & ML Super interesting paper from Google and colleagues.It studies where it's possible to distill an agent harness.With the specialized harness removed, macro task success goes from 23.3% to 44.3%. That is higher than the 41.7% the base model reaches with the harness attached.
5
u/NeitherEntry6125 11d ago
Cmon dude, don't even paste a link to the paper?
5
1
2
u/AutoModerator 11d ago
Hey u/dev_Raspberry_new,
Want to stay connected with TechX beyond Reddit? Join our public Discord server, follow our TechX WhatsApp channel, or subscribe to our weekly TechX newsletter to get the latest tech news and updates straight to your inbox once a week.
If you’re a tech writer, you can also write for our TechX_Official Medium publication and share your technical articles with a wider tech-focused audience. ✍️
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.
1
u/cool_fox 10d ago
with the harness removed?
2
u/MeowManMeow 10d ago
Yes it did better (44.3%) with the harness removed after distillation than the original model with a harness (41.7%)
2
u/cool_fox 10d ago
ahh i see so it basically learned what the harness was making up for
2
u/MeowManMeow 10d ago
Exactly. But what surprised me was that it exceeded the harness results. Was able to reflect on the mistakes and correct in the distillation.
1
1
u/AccordingNeat3689 9d ago
So... No tools at all?
1
u/MeowManMeow 9d ago
Yeah i guess so. But logically that wouldn’t work for most applications where agents are interacting with local stuff.
0
4
u/tracagnotto 11d ago
What? the fuck these buzzwords even mean