Account-Pulse
Analyzes meeting notes to assign a "Health Score" (Green/Yellow/Red).
Compare measured quality, cost, and speed across model tiers. See where a smaller model approaches frontier quality and follow the evidence to an available provider—free, with no API key required.
Showing 214 of 214 tasks · 214 benchmarked
Analyzes meeting notes to assign a "Health Score" (Green/Yellow/Red).
Records why a decision was made to prevent "we already tried that" syndrome later.
Given anonymous employee questions, classify topic, merge duplicates, preserve critical wording, and flag content requiring private follow-up.
Convert Suspicious Activity Report (SAR) data into a narrative report.
Flags transactions that deviate from historical patterns for internal audit.
Given a customer contract, performance obligations, delivery events, and invoices, determine the proposed recognition schedule and flag contract-data conflicts.
Given synthetic business emails, classify likely privilege status, identify attorney involvement and legal-purpose signals, and route uncertain cases for review.
Reviews authentication logic for common flaws (e.g., hardcoded keys).
Given a supplier spreadsheet sample and a target product schema, map source columns, propose type conversions, and identify unmapped required fields.
Given bank transactions and general-ledger entries, match reconciling items and explain outstanding deposits, fees, timing differences, and unexplained variance.
Verify proof-of-claim amounts against the debtor's schedule data.
Debugs and optimizes shell scripts for Linux environment automation.
Creates "Why us vs. them" talking points based on a competitor's URL.
Scans LinkedIn profiles to see if a lead fits the Ideal Customer Profile (ICP).
Given a role competency matrix, generate behavioral questions and anchored scoring criteria that test each required competency without protected-class content.
Reviews performance reviews for gender or racial bias in language.
Extracts usage data (kWh, gallons) and billing cycles from utility statements.
Extracts carrier info, weight, and destination from Bills of Lading.
Given source copy and a measurable style guide, rewrite the content while preserving factual claims and satisfying tone, terminology, and reading-level constraints.
Condenses 10+ articles on the same topic into a single, non-redundant briefing.
Given budget and actual line items with operational drivers, produce a concise variance explanation that correctly attributes amount, direction, and primary cause.
Identifies recurring subscriptions and hidden fees in bank statement exports.
Given a textbook chapter, extract domain terms, source-grounded definitions, and the section in which each term is introduced.
Maps raw bank feeds to the company’s specific Chart of Accounts.
Given a denial reason, coverage policy, and clinician-approved evidence, draft a fact-bound appeal that cites supplied records and avoids invented clinical claims.
Extracts and classifies specific clause types (termination, IP, confidentiality) from a contract.
Given an informed-consent form and a trial-specific checklist, determine whether required risks, contacts, withdrawal rights, and readability constraints are satisfied.
Given synthetic clinician dictation, convert stated information into subjective, objective, assessment, and plan sections without adding diagnoses or facts.
Given a protocol, extract inclusion and exclusion criteria into typed fields while preserving thresholds, timing windows, and conditional logic.
Provides real-time rebuttals for common sales objections (e.g., "Too expensive").
Specialized in translating legacy code (e.g., COBOL, Fortran) to modern languages.
Analyzes existing code for readability and duplication and suggests refactored versions.
Generates production-ready code in the target language from a natural language specification.
Given loan terms and current financial statements, compute covenant ratios, compare them with thresholds, and flag current or near-term breaches.
Scan employee outside activities against company conflict-of-interest policy.
Scans firm-wide client lists to flag potential conflicts of interest for new matters.
Given a product brief, audience, channel limits, and prohibited claims, generate diverse ad variants that satisfy length and compliance rules.
Compare original construction scope to requested change orders.
Given a rubric score and teacher evidence, produce specific growth-oriented feedback that accurately reflects strengths and next steps.
Given synthetic board minutes and a redaction policy, remove personal data and confidential strategy while preserving motions, votes, and publishable governance facts.
Redact PII and sensitive strategy from corporate minutes for public filing.
Given an authorized product record and a marketplace listing, flag suspicious brand, image-description, price, seller, and identifier discrepancies.
Pull dates, witnesses, and exhibits from trial transcripts.
Summarizes CRM account history and deal stage and recent interactions into a one-page account digest.
Sends a "here is what we discussed" email immediately after a sync call.
Given a customer interview transcript and approved facts, produce a problem-solution-result draft without inventing metrics, quotations, or outcomes.
Given synthetic usage, support, renewal, and satisfaction signals, assign a health tier and return the observed factors that drove the classification.
Given issue tickets and merged pull-request summaries, produce release notes grouped by feature, fix, and breaking change while excluding internal-only work.
Assign Harmonized System (HS) codes to product descriptions.
Compares target company profiles against a firm’s investment thesis.
Drafts professional, escalating emails for overdue client invoices.
Flags "TODO" comments and deprecated library usage in the codebase.
Given carrier scans, promised date, weather events, and exception codes, explain the current delay and next expected milestone without inventing an arrival time.
Summarizes multi-hour deposition transcripts into key admissions and contradictions.
Translates a customer's "broken" complaint into a technical JIRA ticket.
Given a patient allergy list, diet order, menu ingredients, and cross-contact rules, classify each meal as allowed, blocked, or requires review.
Confirm IP address and timestamp validity in e-signature audit trails.
Summarizes incoming mail and identifies urgent items needing attention.
Flag attorney-client privilege and work product in email discovery dumps.
Automatically writes documentation (JSDoc, Sphinx) for undocumented code.
Given synthetic documented diagnoses and a supplied code subset, rank candidate codes and cite the exact chart language; always label output as coding-review support.
Given a new issue and an existing issue corpus, rank likely duplicates and distinguish same symptom, same root cause, and merely related requests.
Given a synthetic general ledger and normalization policy, identify candidate nonrecurring expenses, calculate proposed add-backs, and provide the policy basis for each.
Categorizes a massive To-Do list into the Eisenhower Matrix (Urgent/Important).
Given multiple caregiver notes, summarize medications stated as administered, meals, mobility, mood, incidents, and open follow-ups without inferring clinical status.
Given unsubscribe comments, assign one supplied reason label per response and use the fallback for ambiguous or irrelevant feedback.
Rewrites blunt technical answers to be more empathetic and helpful.
Answer employee questions only from supplied current benefits-policy chunks, cite the authoritative section, and abstain when coverage is unspecified.
Given anonymized survey comments, assign supplied themes, quantify prevalence, and extract representative evidence without reidentifying respondents.
Summarize key pollutants and mitigation measures from an Environmental Impact Report.
Suggests a relational database schema based on a description of the data.
Explains cryptic stack traces or server logs in plain English.
Automatically organizes and indexes evidence files into a structured exhibit list.
Given an expense claim and the applicable policy, classify it as approve, reject, or review and return the exact policy rule supporting the decision.
Categorizes raw transactions into specific tax-deductible buckets.
Extracts vendor names, dates, amounts, and VAT numbers from unstructured PDFs.
Given synthetic expense records and interaction notes, classify potential foreign-official risk using a supplied taxonomy and cite the triggering fields.
Distill the Item 19 financial performance representations from a Franchise Disclosure Document.
Drafts personalized emails to users who haven't tried a specific new feature.
Given an asset description, placed-in-service date, jurisdiction, and supplied tax table, assign an asset class and calculate the depreciation schedule.
Converts a spreadsheet of financial forecast data into a plain-English executive narrative.
Suggests how to fill out specific boxes on complex government forms (e.g., SBA, IRS).
Given a data-processing addendum and an Article 28 requirements checklist, classify each requirement as present, missing, or ambiguous with supporting text spans.
Drafts a 90-day success plan based on the customer’s initial goals.
Given structured product attributes and pricing, generate a comparison matrix that accurately distinguishes tier benefits without unsupported superiority claims.
Answer questions only from supplied authoritative contract chunks, cite the supporting chunk, and abstain when the answer is absent or superseded. Available model: neurometric/grounded-document-qa.
Identifies when a customer’s usage patterns suggest they need a higher tier.
Answers employee questions strictly using the internal company handbook.
Turns a resolved support ticket into a draft Knowledge Base article.
Writes personalized cold emails based on a prospect’s recent blog post or tweet.
Prioritizes meeting requests based on the executive’s strategic priorities.
Convert an incident conversation into strict JSON containing final decision, current status, open actions, owners, deadlines, and risks while excluding completed or superseded items. Available model: neurometric/conversation-summary.
Given a job description and an inclusion rubric, identify exclusionary or unnecessarily restrictive language and propose meaning-preserving revisions.
Given a math skill, difficulty, and student interest, generate solvable word problems with verified answers and no irrelevant assumptions.
Given an employee skills profile and open-role requirements, rank eligible roles, identify matched requirements, and surface material gaps.
Given free-form interviewer notes and a competency rubric, convert observations into evidence-linked skill ratings and mark unsupported conclusions.
Summarizes interview transcripts and extracts signals for specific competencies.
Given inventory events, purchase orders, sales velocity, and fulfillment data, identify the primary stock-out cause and generate a customer-safe explanation.
Given historical vendor records and a new invoice, flag suspicious bank-detail changes, duplicate invoice numbers, amount anomalies, and mismatched remittance instructions.
Extract caller-selected invoice fields into exactly typed JSON, preserving source values and returning null for absent fields. Available model: neurometric/document-structured-extraction.
Drafts inclusive, SEO-optimized job descriptions based on raw requirements.
Rewrites blog intros to include high-volume keywords naturally.
Given deployment manifests and CPU-memory utilization summaries, recommend requests and limits within a supplied policy and identify likely throttling or waste.
Suggests specific internal courses based on an employee’s skill gaps.
Improves the Call-to-Action (CTA) effectiveness on landing pages.
Given a customer's delivery note, rewrite it as concise driver steps while preserving access constraints and flagging ambiguous or unsafe instructions.
Given lead attributes, territory rules, and routing history, select the correct owner and explain the exact rules applied in priority order.
Generates first drafts of simple contracts (Service Agreements, LoIs) from basic terms.
Translates complex legislative text into "plain English" for non-legal staff.
Given time-series load-test results, identify saturation point, first failed SLA, likely bottleneck, and the evidence supporting each conclusion.
Categorize lobbying activities into LDA general issue areas for filing.
Given timestamped lab results and reference ranges, identify direction, magnitude, threshold crossings, and missing intervals without offering a diagnosis.
Summarizes why a customer left based on historical communication.
Suggests 5 personalized questions for a 1-on-1 based on the employee's goals.
Given anonymized baseline assessment data and a target skill, draft measurable goals with condition, behavior, criterion, and review period for specialist review.
Build a distribution-rights matrix (territory, media, term) from media contracts.
Drafts pitches to journalists based on their recent coverage areas.
Provides neutral talking points for resolving disputes between team members.
Generates a pre-meeting briefing from CRM contact and company data: goals, risks, and talking points.
Alerts the CSM if a new customer misses their 30-day implementation goals.
Given a question, correct answer, and misconception taxonomy, generate plausible distinct distractors tied to specified misconceptions.
Convert a shopper query into only allowed search filters such as color, fit, material, price range, and occasion while preserving explicit preferences.
Given a customer review and response policy, draft an empathetic reply that acknowledges the issue, avoids prohibited admissions, and offers an approved next step.
Drafts the perfect "ask" for a referral from a happy existing customer.
Acts as a 24/7 assistant for new hires to answer "where do I find" questions.
Ensure non-profit bylaws meet board meeting frequency and quorum requirements.
Drafts offer letters ensuring compliance with local laws and internal bands.
Given a dependency manifest, license texts, and a proposed distribution model, classify compatibility and identify the obligations that drive the result.
Given two OpenAPI specifications, identify removed operations, required-field changes, narrowed enums, response changes, and other client-breaking differences.
Given assignments, rubric results, attendance, and teacher notes, summarize progress, evidence-backed strengths, concerns, and next actions.
Given a patent claim and several candidate prior-art passages, summarize the claimed novelty and the most relevant overlaps without making a legal conclusion.
Given survey comments and a supplied service taxonomy, classify wait time, communication, access, billing, facility, or care-quality concerns.
Given a clinician-approved interpretation and lab result, rewrite it to a target reading level while preserving cautions and escalation instructions.
Explains tax withholdings and deductions to employees in simple terms.
Given performance feedback, flag vague personality judgments, recency bias, unsupported generalizations, and missing outcome evidence.
Given synthetic application logs, identify lines containing possible protected health information, label the PHI type, and avoid flagging safe operational identifiers.
Given application code and a data-handling policy, identify log statements that may expose personal data and suggest policy-compliant structured alternatives.
Checks if a refund request meets the company's terms and conditions.
Identifies coverage limits, deductibles, and effective dates from insurance binders.
Summarizes a pull request diff into a structured review: what changed, why it matters, and risks.
Converts a list of features into a formal Product Requirement Document.
Summarizes the "Holding" and "Ratio Decidendi" of specific court judgments.
Given a coverage policy and synthetic chart, map documented evidence to each requirement and identify unmet or missing documentation without deciding medical necessity.
Compares a new invention description against existing patent abstracts for overlap.
Given a product title, description, and allowed taxonomy, assign supported color, material, style, occasion, and feature tags without inventing attributes.
Given a product requirement, identify undefined actors, missing states, conflicting constraints, untestable language, and unanswered acceptance questions.
Maps resume text into structured JSON fields (skills, years, education) for ATS.
Helps a manager draft a formal promotion justification for an employee.
Given supplied company announcements and public posts, summarize recent buying signals, source each signal, and exclude unsupported personalization.
Refactors slow SQL queries into performant ones for specific databases.
Suggests the best pre-written "macro" based on the user's question.
Generates complex quotes ensuring all discounts and bundles are applied.
Helps managers frame difficult feedback using the "Situation-Behavior-Impact" model.
Calculates and interprets key financial ratios (liquidity, profitability, leverage) from balance sheet data.
Given a passage and target grade level, rewrite it to the requested readability while preserving named entities, facts, and essential vocabulary.
Matches bank statements to internal records and explains discrepancies.
Matches receipts to calendar events to automate expense report narratives.
Compares two contract versions and summarizes what was added, removed, or materially changed.
Converts natural language descriptions into valid, tested Regular Expressions.
Filter out regulatory updates irrelevant to a specific industry.
Generates "What's New" logs for developers from GitHub commit history.
Extracts key dates, rent amounts, and escalation clauses from commercial leases.
Drafts a professional reply to an incoming email based on intent and tone.
Synthesizes multiple web research sources into a structured digest with key findings and gaps.
Analyzes exit interview transcripts for sentiment trends and turnover risks.
Given anonymized order and return history plus policy thresholds, classify a return as standard, review, or high risk and cite the triggering signals.
Identifies power users likely to leave a 5-star review or join a case study.
Given an RFP and a structured response, summarize customer goals, mandatory requirements, differentiators, risks, and unresolved questions in one page.
Analyzes loan applications for qualitative risk factors not seen in numbers.
Given a software master-services agreement and an approved clause policy, identify indemnity and limitation-of-liability deviations and return clause-level suggested edits.
Given a prospect statement and a supplied objection taxonomy, label price, security, timing, authority, competition, or other and identify the decisive phrase.
Given a student profile and supplied scholarship rules, classify eligibility, show each satisfied or unmet criterion, and flag ambiguous evidence.
Generates boilerplate API wrappers given a JSON schema or OpenAPI spec.
Given prior-year and current-year risk-factor sections, identify added, removed, and materially changed risks with paired supporting passages.
Gauge proxy-vote sentiment toward shareholder resolutions.
Drafts public "shoutouts" for wins that align with company values.
Identifies who needs to sign a document based on its type and value.
Extracts product attributes (color, size, material) from supplier catalogs.
Classifies an incoming Slack or Discord message and routes it to the correct channel or team.
Given a textbook passage, generate question-answer flashcards whose answers are fully supported by the source and cover distinct concepts.
Cross-references footnotes in a paper against a database to ensure they are valid.
Given a SQL query, schema, indexes, and execution plan, identify likely bottlenecks and rank safe optimization suggestions without changing query semantics.
Given stack traces with varying request IDs and line noise, cluster incidents by underlying failure signature and provide a representative trace for each cluster.
Given a learning standard, grade, duration, materials, and constraints, produce objectives, activities, checks for understanding, and an exit assessment.
Reviews inventory logs to ensure valuation methods (FIFO/LIFO) are consistent.
Given account history, stakeholders, product usage, opportunities, and risks, produce a one-year plan with evidence-backed goals and next actions.
Fixes layout bugs in CSS based on a description of the desired visual result.
Summarizes a company's weekly activity into an engaging newsletter.
Converts interview notes with a customer into a structured case study.
Translates support tickets while maintaining technical accuracy in the response.
Given a syllabus and institutional checklist, identify present, missing, or conflicting academic-integrity, accessibility, grading, and late-work policies.
Scans product manuals to extract specific warranty durations and conditions.
Given Terraform snippets and a security ruleset, identify overly broad IAM permissions, public storage, exposed networking, and missing encryption with resource locations.
Given opening cash, accounts-receivable aging, accounts-payable schedules, payroll, and recurring commitments, generate weekly inflows, outflows, and ending balances.
Given a purchase order, goods receipt, and invoice, identify quantity, price, tax, and vendor mismatches and recommend the permitted resolution code.
Identifies and archives "Thank you" or automated out-of-office replies.
Suggests folder structures and file naming conventions for cluttered drives.
Reviews content to ensure it matches the specific "voice" of the brand.
Given conversation history and OpenAI-compatible tool definitions, call every required tool with valid supplied arguments or explain missing or invalid parameters without fabricating a call. Available model: neurometric/tool-choice.
Turns a long strategy document into a 2-minute "elevator pitch" for the team.
Given an internal training transcript, extract key rules, required actions, deadlines, owners, and knowledge-check questions.
Suggests relevant memes or cultural trends a brand could participate in.
Summarizes central bank meeting minutes for shifts in "hawkish" vs "dovish" tone.
Guides users through a step-by-step interactive troubleshooting flow.
Given a function contract and implementation, generate executable unit-test cases covering happy paths, boundaries, invalid inputs, and observed failure branches.
Explains the technical steps needed to move from an old version to a new one.
Ranks support tickets by sentiment and severity for immediate action.
Given vendor-master records, identify likely duplicates across name, address, tax ID, and payment details while distinguishing related but separate legal entities.
Given a security-researcher report, extract affected asset, prerequisites, reproduction steps, impact claim, evidence, and missing details for engineering triage.
Given safety logs, classify incident type and contributing factors, merge duplicates, and summarize recurring prevention opportunities.
Alerts a manager if a high-value customer expresses any dissatisfaction.
Given a dealer application, extract legal entity, principals, references, requested limit, tax identifiers, and missing fields into typed JSON.
Given a current schema, target schema, traffic pattern, and deployment constraints, order expand-migrate-contract steps and flag unsafe operations.
Every comparison uses the same task-specific process. Results are directional rather than a production SLA, and each task shows its sample, rubric, judge, limitations, and pricing provenance.
We create and freeze a representative set of real task cases, so every model is graded on exactly the same work.
A task-specific rubric spells out what counts as correct — and what counts as a failure — before any model runs.
Qualified candidates from all four tiers run on the frozen holdout and are scored by an LLM judge for comparable pass rates.
Measured token profiles meet the latest available provider prices, with source and freshness shown so every estimate can be checked.