This report documents a benchmark of large language models on InterSystems ObjectScript and IRIS platform tasks. Model responses to 140 hand-written items are graded by compiling and executing them ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results