Does the latter one look like an function that response “Safe|Dangerous|Unknown” and is never wrong about Safe or Dangerous judgements, but might rarely fall back to the oracle that always answers Safe|Dangerous? If that’s the case that seems quite valuable even in the absence of the perfect oracle, if you’ve got something which can categorize between “definitely safe” and “might not be safe” and never judges something as “definitely safe” when it’s actually dangerous, and usually judges safe things as safe (i.e. it’s not just a rock with “might be unsafe” written on it), that seems like it gets you most of the value already.
Does the latter one look like an function that response “Safe|Dangerous|Unknown” and is never wrong about Safe or Dangerous judgements, but might rarely fall back to the oracle that always answers Safe|Dangerous? If that’s the case that seems quite valuable even in the absence of the perfect oracle, if you’ve got something which can categorize between “definitely safe” and “might not be safe” and never judges something as “definitely safe” when it’s actually dangerous, and usually judges safe things as safe (i.e. it’s not just a rock with “might be unsafe” written on it), that seems like it gets you most of the value already.
the latter is safe|dangerous|unknown.