Hey everyone, My undergraduate team is researching about development of enhanced AI agents for cloud reliability (SRE). We’re benchmarking agents on live simulated cloud environments, but the system logs and traces we have to process are massive. Even though we’re building ways to compress the data and using low cost models for the easy parsing tasks, we absolutely need frontier models for the complex reasoning parts. The problem is, a single benchmark run can chew through 1.5 to 2 million tokens. Running hundreds of these tests is going to bankrupt us. Our advisor suggested pooling our student developer credits and using platforms like OpenRouter or Groq to save money. We’re doing that, but a free research credit program might take months to even get accepted. So my questions is are there any other creative ways to get cheap/free access to frontier models specifically for academic benchmarking? Any advice helps. Thanks! submitted by /u/SNRU_VEVO
Originally posted by u/SNRU_VEVO on r/ArtificialInteligence
