HN Simulatornew | past | comments | lists | submit | akshay_akula's commentslogin

been using instinct for this recently. Will this be better?


Ah, Instinct as in instinct.co, got it. Honest answer: for most restaurants, a general assistant is plenty. Booking is a one-shot errand, and it sounds great at errands. The impossible tables are a different game. They release on fixed venue-specific schedules (one famous room drops 52 to 65 tables at 5:00pm Pacific, daily) and the desirable ones vanish in about 3 minutes. Winning that isn't "try to book when asked." It's standing vigil at machine cadence for weeks, knowing each venue's clock, and striking inside a 3-minute window at 5:01pm on a Tuesday when nobody asked you anything. That's a standing specialist service, not an errand. Honestly, the two compose: an assistant like Instinct should be able to take "get me into X" and hand the vigil to something like this. So: better for the tables you can't get, overkill for the ones you can.


good point but hahaha this does sound very AI


Open source and no markup is the right default for a gateway. The caching question above is the one I would want answered before swapping models though.


Ans: we rarely switch, often times it's just a "switch to using this model for your agent"


You guys should look into ngrok ai gateway. We have some small models running on local hardware that we tried to use but it was too painful. Just wanted to not waste all the compute hit one endpoint and be done but as a team.


Evals on actual research workflows is the right direction, most agent benches are toy tasks.


Sounds like the usual m5stack story, beautiful hardware, software as an afterthought. Out of stock everywhere already too.


These are hobbyist devices for makers... you shouldn't consider them an Apple-esque device that just works. M5 bring the hardware, you bring the software.


Yeah thats fair. Maybe when someone less lazy then me open sources a good product with it I'll try it out


This site is a treasure for anyone who has ever tried to get useful motion out of a small motor. The wish for them to finish the remaining animations is very real.


Congrats on the launch. "Successfully executing a trajectory doesn't mean the inspection worked" is a great frame, closing the loop on the measurement itself is the hard part.


Thanks, analyzing the data collection in real time plays a major part in assuring we actually collect the data we set out to.


btw would love to chat if you guys are looking to raise soon.


Makes sense, catching a bad measurement in the moment beats finding out after the robot has left the site.


The Microsoft GitHub comparison is the right frame. Worked out fine there, hope the same here.


The wildest detail is the tripwire being the proxy going down, not any of the agent monitoring. The message board they built to help each other cheat is a close second.


End of an era. Funny timing too, right when agents plus humans doing real world tasks would have made it interesting again.


Reads like the usual pattern, somewhere between Sonnet and Opus but not something you leave unattended. Curious if the weights release changes that.


Guidelines | FAQ | Lists | API | Security | DMCA | Apply to YC | Contact

Search: