Explore Problems
Showing 1,354 of 9,941 problems · matching your filters
No reliable lightweight method to evaluate whether AI prompt tweaks actually improve outcomes
Developers modifying AI prompts or workflows rely on intuition rather than systematic evaluation, making it hard to know if changes genuinely improve performance. The lack of simple evaluation frameworks causes regressions to go undetected. A growing problem as AI-assisted workflows become standard in software development.
Inconsistent Quality and Missing Features When Serving Open-Weight Vision Models
Developers running open-weight VLMs, OCR models, and vision transformers in production struggle with undocumented quantization differences that silently degrade OCR and spatial accuracy, poor video input support across most providers, and the operational complexity of building document-inference pipelines. These issues make it hard to trust a model listing's claimed capabilities or achieve consistent output quality at scale.
Online Used-Car Sellers' Inspections Miss Defects Independent Mechanics Catch Instantly
Buyers of online used cars trust the seller's multi-point inspection report, but discover significant defects that multiple independent mechanics identify within seconds. Sellers dispute responsibility for the hidden defect and only partially reimburse repair costs, leaving buyers to cover the difference on a large purchase.
Undisclosed Vehicle Defects Despite Dealer's Pre-Purchase Inspection Claims
Used car buyers rely on a dealer's stated multi-point inspection to catch pre-existing defects, but discover mechanical faults only after purchase. Buyers have no independent way to verify inspection records, leaving them to bear repair costs after warranty support is denied.
AI Answer Engines Cite Competitors Instead of a Business's Own Site
A founder's brand-visibility report revealed that AI chat and search tools were citing her competitors' homepages rather than her own when responding to relevant queries, even though she was listed on review sites. The finding points to a growing gap between traditional SEO presence and how generative AI engines choose which sources to surface.
Lack of Trustworthy Stop Criteria When Delegating Tasks to AI Agents
When delegating work to an autonomous AI agent, users struggle to define what evidence would let them confidently trust that the agent has reached a correct stopping point. This is a systemic gap in how agent output is verified before a human accepts it, rather than a flaw in any single tool.
Multistate Employee Tax Compliance Gaps in Payroll Software
Small and mid-sized businesses struggle to navigate multistate payroll tax requirements when employees work across state lines. Payroll platforms like Gusto provide insufficient guidance on which forms to file, when tax nexus applies, and which employees qualify for exemptions. This creates compliance risk and administrative burden for HR teams.
AI Coding Agents Lack Access to Production Runtime Context During Debugging
AI coding agents operate without real-time production telemetry, forcing them to debug blindly using sampled or delayed observability data. Development teams face review fatigue from deduplicated and incomplete signals when agents attempt automated fixes. Bridging the gap between agent context and production-level runtime data is an emerging need as AI-assisted development matures.
Insurance Companies Report Customers to Credit Bureaus Without Adequate Dispute Process
Consumers who switch insurers before policy expiry are at risk of being reported to credit bureaus by their former insurer for refusing overlap charges. The lack of a standardized grace period or dispute pathway leaves customers with damaged credit and no clear recourse. This gap between insurance billing practices and credit reporting consequences is a structural consumer protection failure.
Building Durable Long-Running Tasks Requires Manual Infrastructure
Developers building agent loops, ETL pipelines, and billing workflows must wire together queues, worker pools, retry logic, and state management themselves — infrastructure that doesn't differentiate their product. The operational overhead scales with reliability requirements, making correctness expensive.
Slack global search returns irrelevant results and huddles quality degraded
User reports Slack global search returns poor matches with unclear filtering, and huddles feature quality has regressed to the point of switching to Google Meet. More detailed review confirming search and real-time communication regressions.
Zero-Knowledge Proof Generation Is Too Slow and Memory-Intensive for Mobile Applications
Generating zero-knowledge proofs on mobile devices requires prohibitive compute time and RAM, making privacy-preserving mobile applications impractical at current performance levels. The gap between ZK proof requirements and mobile hardware constraints is a structural barrier to building privacy-first mobile products. As privacy regulation grows and user expectations rise, this bottleneck blocks an entire class of applications from being built.
Insurers dispute independent roofing assessments to minimize storm-damage payouts
A homeowner's storm-damaged roof was assessed by a State Farm adjuster as repairable with a patch, but two independent professional roofers determined a full replacement was required per manufacturer and industry standards. State Farm declined to share the adjuster's report or offer mediation, and the poster ties the dispute to a documented nationwide industry pattern of minimizing roof-damage payouts.
Payday Lenders Contact Employer Despite Explicit Verbal Cease Requests
Sunset Finance repeatedly contacted a consumer's employer after being told to stop, violating FDCPA harassment prohibitions. Payday lenders use workplace contact as a coercive collection tactic, causing reputational damage at the consumer's job.
Auto Lenders Charge Late Fees Despite Confirmed Written Payment Arrangements
Credit Acceptance charged late fees on dates that were part of a documented payment arrangement, confirmed in writing via email and text. The lender's billing system ignored the agreed arrangement, creating fees despite customer compliance.
Nutrition Tracking Abandonment Driven by Barcode Scanning and Manual Calorie Logging
Traditional nutrition apps require users to scan barcodes or manually search and log every food item, creating enough friction to cause habitual abandonment. The effort-to-insight ratio is poor: extensive data entry yields delayed nutritional feedback. This behavioral barrier prevents consistent tracking even among users who understand the health value of monitoring their diet.
Mortgage Servicers Fabricating Missed Payments After Hardship Recovery
Mortgage servicers falsely claim payments were missed during hardship periods despite consumer records showing all payments were made. Fabricated delinquencies trigger fee assessments and negative credit reporting that compound the harm of the original hardship. Consumers who document their payments still cannot force servicers to correct fraudulent delinquency records.
Mortgage Servicers Denying Permanent Modifications After Trial Plan Completion
Homeowners who successfully complete trial loan modification plans are denied permanent modifications, often without explanation. This pattern traps consumers in limbo after fulfilling all required trial period payments. The lack of automatic conversion from trial to permanent modification when trial criteria are met is a well-documented servicer abuse pattern.
Credit Bureaus Reinsert Blocked Identity-Theft Accounts Without Verification
Identity theft victims who file a police report and formally request a fraud account block under FCRA Section 605B find the fraudulent account reinserted onto their credit report without new verifiable evidence. Credit bureaus refuse to re-investigate, misdirect victims to the creditor, and fail to follow legally mandated dispute procedures, blocking victims' access to financing.
Debt Collector Falsely Claims Debt Ownership to Credit Bureaus in FCRA Violation
A debt collector falsely represents to credit reporting agencies that it owns a debt, resulting in inaccurate credit report entries. FCRA violations from false ownership claims damage consumer credit without legal basis. Enforcement gaps allow collectors to report debts they do not legitimately own.