Skip to content
SIH Buddyby Ganeev Singh
Dev

๐Ÿ”ฅ Roast My Pick ยท SIH26063

Integrated Polar Science Outreach, Knowledge Repository and Media Dissemination Portal

Ministry of Earth Sciences (MoES)

Brutal72/100

Bold. Let us find out precisely how bold, in the order a panel will find out.

Proceed with caution. The corpus is real and the build is safe, but an archive with an AI post generator has a very low ceiling โ€” take it only if you will make grounded, citation-locked generation and multimodal search the substance rather than shipping a CMS. Roughly 95โ€“220 teams are expected to go here.

The receipts

Every red flag on this statement, in full. These are the four places it bites.

  1. Exhibit A

    An archive plus a text generator is among the least technically distinctive products you can build, and a panel will have seen several retrieval-and-generate submissions before yours

  2. It gets worse

    Generated science communication that subtly misstates a finding is worse than no post at all, so the editorial review queue is a requirement rather than a nicety and should be visible in your demo

  3. Still reading?

    Six media types is deceptively broad โ€” video ingestion, transcription and indexing alone can consume the whole event if you let it

  4. And the finisher

    There is no measure of good outreach in the statement, so any claim that your generated content is effective rests on your own judgement and a judge may simply disagree

The damage report

Every score this statement earned, and what each one actually costs you.

  • Feasibility

    4/5

    Actually buildable, which on this slate is rarer than it sounds. Do not squander it on scope.

    Indian Antarctic expedition reports, station documentation and imagery are published openly enough to assemble a genuine corpus, and repository search plus grounded text generation are mature and well-supported โ€” nothing here depends on data you cannot obtain.

  • Innovation scope

    4/5

    There is something genuinely new here. Do not bury it under another dashboard.

    At 213 characters the statement names what the portal holds and that it should generate content, and prescribes nothing about how โ€” the retrieval design, the generation grounding and the editorial workflow are all yours.

  • Clarity

    2/5

    Nobody is sure what is being asked, quite possibly including the people who asked it.

    Short, but unlike the neighbouring polar statements it does at least enumerate the content types and identify two distinct functions in archiving and generating, so you are not starting from nothing โ€” though there is no audience, no scale and no measure of what good output would be.

  • Acceptance potential

    2/5

    The numbers do not like you. Bring something the numbers cannot see.

    A content repository with an AI caption generator is close to the lowest-ceiling software shape available โ€” the archive half is a CMS and the generation half is an LLM call, so there is very little for a judge to reward beyond execution and the outreach framing does not create technical substance.

  • Effort

    Heavy

    Heavy. Somebody on this team is not sleeping in week three. Pick who, on purpose.

    Ingestion across six media types, metadata extraction, a search layer, a generation pipeline, an editorial review queue and a scheduling interface is six components, and handling video and imagery properly is more work than the text side.

  • Demo-ability

    Medium

    Demoable, if you rehearse it. Nobody rehearses it.

    Search over a real archive works reliably and grounded generation with highlighted sources is a decent moment, but a content portal generating social posts is a familiar sight in 2026 and nothing here will surprise a panel.

  • Data

    None supplied

    No dataset comes with this one, so every accuracy figure you quote is a number about labels you invented.

    Nothing is provided with the statement. You are sourcing, cleaning and labelling it yourself, and that work is invisible in the demo but very visible in the questions.

The demo they will have already seen

Somewhere around 95โ€“220 teams are heading here, and the description is doing the choosing for most of them. They will read the same brief, reach the same architecture, and build a version of the same demo you are planning. Being correct is the floor. If your five minutes could be swapped with the team before you and nobody in the room would notice, you have not picked badly โ€” you have built predictably, which costs exactly the same and hurts more.

What survives

The ground worth standing on when the questions start.

  • Indian polar expedition material is genuinely published and downloadable, so unlike the neighbouring NCPOR statements your corpus is real rather than invented
  • Grounding generated posts to specific passages in a source report is a real and demonstrable discipline, and a portal that refuses to write anything it cannot cite is a defensible differentiator over freeform generation
  • Semantic search across photographs and video, not just text, is the harder and more interesting half and almost no competing team will attempt it

None of that means do not pick it. It means do not walk into that room having heard any of this for the first time from a judge.

The framing is a joke. The findings are not โ€” they are the same analysis on the statement page, and every line above is attached to a score or a fact in the record. It is one opinion with its reasoning attached, so argue with it before you trust it.