Remix clone Hacker News

new | show | ask | jobs Github

	▲	empath75 5 hours ago
		I am not sure we are at the "efficiency" phase of this. Even if you just wire this output (or probably multiples running different counterfactuals) into a multimodal LLM that interprets the video and uses it to make decisions, you have something new.