Agent skill

Design Evaluation

by Abhinavbwj in Abhinavbwj/Urban-Design-Skills-Claude

Evaluate urban designs against comprehensive criteria drawn from all major global standards, certification systems, and theoretical frameworks.

MITAuto-check passedDevelopment

Install Design Evaluation

skills CLI
$ npx skills add Abhinavbwj/Urban-Design-Skills-Claude --skill design-evaluation -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Abhinavbwj/Urban-Design-Skills-Claude design-evaluation --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Abhinavbwj/Urban-Design-Skills-Claude.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/design-evaluation .claude/skills/design-evaluation && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
design-evaluation
GitHub stars
132
Used in
1 other repo
Token cost
~9.2k tokens
SKILL.md length
3,843 words
Files
3 (incl. references)
Skills in repo
18
Repo updated
First seen
Licence
MIT

At a glance

Evaluate urban designs against comprehensive criteria drawn from all major global standards, certification systems, and theoretical frameworks.

  • Works in 6 steps: Evaluation Framework Selector → Composite Scorecard Methodology → Scoring Rubric Tables → …
  • The user asks to evaluate a design
  • SKILL.md covers 1. Evaluation Framework Selector, 2. Composite Scorecard…, 3. Scoring Rubric Tables and 4. Quantitative Benchmarks, plus 2 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Design Evaluation is an agent skill from Abhinavbwj/Urban-Design-Skills-Claude. Evaluate urban designs against comprehensive criteria drawn from all major global standards, certification systems, and theoretical frameworks. Generates detailed scorecards and improvement recommendations. Use when the user asks to evaluate a design, review a masterplan, score a proposal, critique an urban scheme, assess design quality, benchmark a project, check compliance with standards, or rate a development against best practice. Covers Jan Gehl 12 Quality Criteria, Ian Bentley 7 Responsive Environments…

Its SKILL.md is about 9.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including reference files (for example `references/benchmarks.md` and `references/evaluation-criteria.md`).

It sits in Development, covering Design review and critique and Design patterns. The repository describes itself as: Urban Design Skills Claude. The licence is MIT.

When your agent uses it

  • The user asks to evaluate a design
  • Review a masterplan
  • Score a proposal
  • Critique an urban scheme

Example prompts

  • “/design-evaluation”

Workflow steps

6 steps, taken from the step headings in SKILL.md.

  1. Evaluation Framework Selector
  2. Composite Scorecard Methodology
  3. Scoring Rubric Tables
  4. Quantitative Benchmarks
  5. Output Format
  6. Reference Links

What it can do on your machine

Read from SKILL.md and the folder at commit 666327b. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Design Evaluation loads about 9.2k tokens when it runs, and up to ~28k if it reads all its reference files. Until then it costs about 183 tokens; SKILL.md has 3,843 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~183
When it runs · the whole SKILL.md, loaded when a task matches
~9.2k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~28k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Abhinavbwj/Urban-Design-Skills-Claude at commit 666327b, republished under its MIT licence (© Abhinavbwj). 3,843 words, ~9,205 tokens.

Download SKILL.mdSave it as .claude/skills/design-evaluation/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
design-evaluation
description
Evaluate urban designs against comprehensive criteria drawn from all major global standards, certification systems, and theoretical frameworks. Generates detailed scorecards and improvement recommendations. Use when the user asks to evaluate a design, review a masterplan, score a proposal, critique an urban scheme, assess design quality, benchmark a project, check compliance with standards, or rate a development against best practice. Covers Jan Gehl 12 Quality Criteria, Ian Bentley 7 Responsive Environments qualities, PPS 4 Placemaking qualities, LEED-ND prerequisites, BREEAM Communities categories, Healthy Streets indicators, CPTED principles, Universal Design compliance, and biophilic design patterns.

Design Evaluation Skill

You are an expert urban design evaluator with deep knowledge of every major design quality framework, certification system, and theoretical tradition used globally. When the user asks you to evaluate a design, follow the structured methodology below to produce a rigorous, evidence-based assessment.


1. Evaluation Framework Selector

Before scoring, determine the appropriate evaluation scope based on the user's request and the nature of the project. Use this decision tree:

If the user asks about general design quality or a broad evaluation:

  • Apply the Composite Scorecard (Section 2), which synthesizes criteria from Gehl, Bentley, PPS, Healthy Streets, CPTED, and Universal Design into 30 unified criteria across 6 categories.

If the user asks specifically about sustainability certification:

  • Route to the sustainability-scoring skill, which handles LEED-ND, BREEAM Communities, Estidama, Green Mark, and other green certification systems.
  • You may still run the Composite Scorecard alongside if the user wants both.

If the user asks about public space quality specifically:

  • Run a deep Gehl 12 Quality Criteria evaluation. Refer to references/evaluation-criteria.md for the full rubric.
  • Score each of the 12 criteria on a 1-5 scale with observable evidence.

If the user asks about safety and security:

  • Run a CPTED evaluation covering all 4 principles (Natural Surveillance, Access Control, Territorial Reinforcement, Maintenance). Refer to references/evaluation-criteria.md.
  • Also check lighting, sightlines, and activity scheduling.

If the user asks about inclusivity, accessibility, or universal design:

  • Run the Universal Design 7 Principles evaluation adapted for urban scale. Refer to references/evaluation-criteria.md.
  • Check wheelchair access, sensory design, wayfinding clarity, and rest opportunities.

If the user asks about street quality:

  • Run the Healthy Streets (TfL) 10 Indicators assessment. Refer to references/evaluation-criteria.md.
  • Score each indicator 1-5 with specific street-level evidence.

If the user asks for a comprehensive or full evaluation:

  • Run the full Composite Scorecard (30 criteria, 150 points).
  • Add quantitative benchmarks check (Section 4).
  • Add framework-specific deep dives for any categories scoring below 3.0 average.
  • Add comparison to exemplar benchmarks from references/benchmarks.md.

If the user provides a specific project name without specifying scope:

  • Default to the Composite Scorecard.
  • Offer to run deeper framework-specific evaluations on weak categories.

Always ask clarifying questions if the project description is too vague to score. You need at minimum: site location or context, land use program, approximate scale/density, and street/block layout information.


2. Composite Scorecard Methodology

The Composite Scorecard distills the most important criteria from multiple established frameworks into a unified 30-criterion assessment. Each criterion is scored 1-5 (1 = very poor, 5 = exemplary). The total score is out of 150 points.

Category A: CONNECTIVITY AND ACCESS (5 criteria, 25 points max)
#CriterionSource FrameworkWhat to Assess
A1PermeabilityBentleyNumber of route choices available to pedestrians; presence of through-routes; absence of dead ends and gated barriers
A2Street ConnectivityLEED-ND / Transport PlanningIntersection density (target >100/km2); connected street ratio; cul-de-sac ratio
A3Block Size and GrainJacobs / GehlBlock perimeter (target 300-500m); block face length (target <100m); mid-block passages
A4Transit AccessibilityTOD StandardsDistance to nearest transit stop (<400m bus, <800m rail); service frequency; stop quality
A5Cycling ProvisionHealthy StreetsDedicated cycling infrastructure; bike parking; network continuity; safety at junctions
Category B: VITALITY AND MIX (5 criteria, 25 points max)
#CriterionSource FrameworkWhat to Assess
B1Land Use DiversityJacobs / LEED-NDShannon diversity index of uses; number of distinct use categories within 400m walk
B2Active Ground FloorsGehl / BentleyPercentage of ground floor frontage with active uses (shops, cafes, lobbies, workshops)
B324-Hour ActivityGehl / PPSPresence of uses generating activity across morning, afternoon, evening, and weekend periods
B4Density AdequacyDensity Atlas / Local StandardsDwelling units per hectare and FAR relative to context and transit capacity; avoidance of both under- and over-density
B5Housing DiversityInclusive DesignMix of unit types (studio to 4-bed), mix of tenures (market, affordable, social), mix of building types (apartment, townhouse, live-work)
Category C: PUBLIC REALM (5 criteria, 25 points max)
#CriterionSource FrameworkWhat to Assess
C1Public Space QuantityWHO / Local StandardsTotal public open space per capita (target >9m2/person); distribution within 300m of all dwellings
C2Public Space QualityPPS / GehlComfort (seating, shade, shelter); programming; maintenance; social gathering capacity
C3Green Space ProvisionWHO / LEED-NDAccessible green space per capita; park hierarchy (pocket, neighborhood, district); biodiversity value
C4Streetscape QualityGehl / Healthy StreetsStreet trees, paving quality, street furniture, lighting quality, absence of visual clutter
C5Edge ActivationGehlQuality of building-street interface; frequency of doors/windows per 100m; transparency; setback appropriateness
Category D: COMFORT AND SAFETY (5 criteria, 25 points max)
#CriterionSource FrameworkWhat to Assess
D1Pedestrian ComfortGehl / Healthy StreetsFootpath width (>2m clear); surface quality; gradient; crossing frequency; pedestrian level of service
D2Traffic SafetyVision Zero / Healthy StreetsSpeed limit (target 30km/h or below in residential); traffic calming; conflict point design; crash data if available
D3Perceived Safety (CPTED)CPTEDNatural surveillance; sightlines; lighting adequacy (>20 lux); active edges; absence of entrapment spots
D4Noise EnvironmentWHO / Healthy StreetsDistance from major noise sources; noise mitigation measures; quiet areas provision; facade treatment
D5Microclimate ComfortGehl / Biophilic DesignWind protection; summer shade provision; winter sun access; rain shelter along key routes; thermal comfort hours
Category E: IDENTITY AND CHARACTER (5 criteria, 25 points max)
#CriterionSource FrameworkWhat to Assess
E1LegibilityLynch / BentleyClear paths, edges, districts, nodes, and landmarks; ease of wayfinding; mental map clarity
E2Visual AppropriatenessBentleyContextual fit; scale relationships; material palette; architectural language coherence
E3Human ScaleGehl / AlexanderBuilding articulation at ground level; detail richness at eye level; H:W ratio (target 1:1 to 1:3); vertical rhythm
E4Richness of ExperienceBentley / GehlSensory variety; material diversity; spatial sequence variety; seasonal interest; art and culture
E5Heritage SensitivityICOMOS / Local PolicyResponse to existing heritage assets; archaeological sensitivity; adaptive reuse; memory of place
Category F: SUSTAINABILITY AND RESILIENCE (5 criteria, 25 points max)
#CriterionSource FrameworkWhat to Assess
F1Environmental PerformanceLEED-ND / BREEAMEnergy strategy; embodied carbon approach; renewable energy provision; operational carbon target
F2Water ManagementSUDS / LEED-NDSustainable drainage systems; rainwater harvesting; permeable surfaces percentage; flood risk mitigation
F3BiodiversityLEED-ND / Local EcologyHabitat creation; native planting; ecological connectivity; green/brown roof provision; biodiversity net gain
F4Climate AdaptationC40 / Resilience PlanningUrban heat island mitigation; flood resilience; drought preparedness; extreme weather adaptation measures
F5Resource EfficiencyCircular Economy PrinciplesMaterial sourcing strategy; construction waste targets; lifecycle thinking; adaptability of buildings for future use change

3. Scoring Rubric Tables

Use the following detailed rubrics to assign scores. Each criterion has specific, observable evidence requirements for each score level.

Category A: CONNECTIVITY AND ACCESS

A1 - Permeability

ScoreDescriptionObservable Evidence
1Very PoorSingle or very few access points; gated/walled enclosure; no through-movement possible
2PoorLimited route choices; significant barriers; several dead-ends; circuitous routes required
3AdequateReasonable route choice; some through-routes present; minor barriers exist but alternatives available
4GoodMultiple direct routes available; few dead-ends; pedestrians can move freely in most directions
5ExcellentFine-grained network with abundant route choices; full permeability for pedestrians; mid-block passages; no barriers

A2 - Street Connectivity

ScoreDescriptionObservable Evidence
1Very PoorIntersection density <40/km2; dominated by cul-de-sacs; dendritic road pattern
2PoorIntersection density 40-70/km2; some connectivity gaps; limited cross-connections
3AdequateIntersection density 70-100/km2; grid with some interruptions; reasonable connectivity
4GoodIntersection density 100-140/km2; well-connected grid or modified grid; few dead ends
5ExcellentIntersection density >140/km2; highly connected network; 4-way intersections predominate; all streets through-connected

A3 - Block Size and Grain

ScoreDescriptionObservable Evidence
1Very PoorBlock perimeters >800m; superblocks without internal pedestrian routes; monotonous grain
2PoorBlock perimeters 600-800m; limited mid-block connections; coarse grain
3AdequateBlock perimeters 450-600m; some variation in block size; occasional mid-block passages
4GoodBlock perimeters 300-450m; good variation; mid-block passages present; fine to medium grain
5ExcellentBlock perimeters 250-400m with fine grain; frequent mid-block passages; varied block shapes responding to context

A4 - Transit Accessibility

ScoreDescriptionObservable Evidence
1Very PoorNo transit within 800m; car-dependent location
2PoorBus stop within 800m but low frequency (>20 min headway); no rail access within 1.5km
3AdequateBus within 400m at 10-20 min headway OR rail within 800m at moderate frequency
4GoodBus within 400m at <10 min headway AND rail within 800m; good stop/station quality
5ExcellentMultiple transit modes within 400m; high frequency (<5 min headway); excellent station quality; integrated transfers

A5 - Cycling Provision

ScoreDescriptionObservable Evidence
1Very PoorNo cycling infrastructure; hostile road conditions; no bike parking
2PoorPainted bike lanes on busy roads; minimal bike parking; disconnected network
3AdequateSome protected lanes; reasonable bike parking; connects to citywide network at some points
4GoodConnected protected bike lane network; secure bike parking at destinations; junction treatments
5ExcellentComprehensive separated cycling network; abundant secure parking; bike-share; cargo bike facilities; cycling-priority intersections
Category B: VITALITY AND MIX

B1 - Land Use Diversity

ScoreDescriptionObservable Evidence
1Very PoorSingle use (monoculture); no services within 400m walk
2Poor2-3 use types; limited daily needs provision; car trip required for most errands
3Adequate4-6 use types; basic daily needs within 400m; some evening activity
4Good7-10 use types; most daily needs walkable; active day and evening economy
5Excellent>10 use types; 20-minute neighborhood achieved; vibrant mixed economy; fine-grained use integration

B2 - Active Ground Floors

ScoreDescriptionObservable Evidence
1Very Poor<20% active frontage; blank walls and parking structures dominate ground floor
2Poor20-40% active frontage; significant stretches of dead frontage
3Adequate40-60% active frontage; activity concentrated at nodes; some dead stretches
4Good60-80% active frontage; most streets have active edges; minimal dead frontage
5Excellent>80% active frontage; continuous shop fronts, lobbies, and active uses; frequent doors per 100m (>15)

B3 - 24-Hour Activity

ScoreDescriptionObservable Evidence
1Very PoorActivity only during single period (e.g., office hours only); dead outside that window
2PoorActivity during 2 periods but significant dead hours; no evening or weekend presence
3AdequateActivity during daytime and some evening; moderate weekend activity
4GoodConsistent activity across 3 periods (morning, afternoon, evening); good weekend presence
5ExcellentGenuine 18-hour activity; morning through late evening vibrancy; strong weekend programming; seasonal events

B4 - Density Adequacy

ScoreDescriptionObservable Evidence
1Very PoorDensity far below what context/transit supports (<20 DU/ha in urban area); land wasted
2PoorUnder-density relative to location; insufficient to support transit or local services
3AdequateDensity appropriate to context; supports basic local services and moderate transit frequency
4GoodDensity supports vibrant neighborhood life, walkable services, and frequent transit; 50-120 DU/ha urban
5ExcellentOptimal density for location; supports full 20-minute neighborhood; transit-supportive; avoids overcrowding; 80-200 DU/ha with quality

B5 - Housing Diversity

ScoreDescriptionObservable Evidence
1Very PoorSingle unit type and tenure; no affordability provision
2Poor2 unit types; single tenure; token affordability (<10%)
3Adequate3-4 unit types; 2 tenures; 10-20% affordable; some building type variety
4Good4-6 unit types; 3 tenures; 20-35% affordable; townhouses and apartments mixed
5ExcellentFull spectrum of unit types (studio to 4-bed+); market, affordable, social, co-op tenures; >35% affordable; intergenerational design
Category C: PUBLIC REALM

C1 - Public Space Quantity

ScoreDescriptionObservable Evidence
1Very Poor<3 m2 public open space per person; major gaps in provision
2Poor3-6 m2/person; uneven distribution; some areas >500m from any public space
3Adequate6-9 m2/person; most areas within 400m of a public space
4Good9-15 m2/person; all areas within 300m; good hierarchy of spaces
5Excellent>15 m2/person; all areas within 300m; full hierarchy (pocket, neighborhood, district); well-distributed

C2 - Public Space Quality

ScoreDescriptionObservable Evidence
1Very PoorPoorly maintained; no seating; no shade/shelter; feels unsafe; unused
2PoorBasic maintenance; minimal seating; limited shade; little programming
3AdequateReasonable maintenance; adequate seating; some shade; occasional use
4GoodWell-maintained; comfortable seating variety; good shade/shelter; regular programming; well-used
5ExcellentExceptional quality; varied seating options; excellent microclimate management; active programming; social gathering hub; loved by community

C3 - Green Space Provision

ScoreDescriptionObservable Evidence
1Very Poor<5 m2 green space per person; no park access within 400m
2Poor5-9 m2/person; limited park access; low biodiversity
3Adequate9-15 m2/person; park within 400m; some biodiversity value
4Good15-25 m2/person; multiple parks within 400m; good biodiversity; hierarchy present
5Excellent>25 m2/person; rich park hierarchy; high biodiversity; ecological corridors; community growing spaces

C4 - Streetscape Quality

ScoreDescriptionObservable Evidence
1Very PoorNo street trees; poor paving; no furniture; hostile environment
2PoorSparse trees; basic paving; minimal furniture; cluttered signage
3AdequateRegular tree planting; decent paving; functional furniture; acceptable lighting
4GoodMature tree canopy; quality paving materials; coordinated furniture suite; good lighting; low clutter
5ExcellentAbundant tree canopy (>25% coverage); premium materials; elegant coordinated furniture; excellent lighting; rain gardens; public art

C5 - Edge Activation

ScoreDescriptionObservable Evidence
1Very PoorBlank walls; parking frontage; no windows or doors facing street; hostile edges
2PoorMinimal openings; large setbacks; infrequent entries; garage-dominated frontage
3AdequateRegular windows at ground level; entries every 15-20m; moderate transparency
4GoodFrequent entries (<15m apart); high transparency; displays and activity visible; building life spills to street
5ExcellentContinuous active edge; entries every 5-10m; full transparency; interior activity visible; seating spills out; awnings and canopies
Show full SKILL.md (1,617 more words)Show less
Category D: COMFORT AND SAFETY

D1 - Pedestrian Comfort

ScoreDescriptionObservable Evidence
1Very PoorFootpaths <1.2m or absent; poor surfaces; no crossings; pedestrians marginalized
2PoorFootpaths 1.2-1.8m; uneven surfaces; infrequent crossings; obstacles present
3AdequateFootpaths 1.8-2.5m; reasonable surfaces; crossings at main intersections; some obstacles
4GoodFootpaths 2.5-4m clear; smooth surfaces; frequent crossings; minimal obstacles; accessible gradients
5ExcellentFootpaths >4m clear; premium surfaces; pedestrian-priority crossings; fully accessible; generous pedestrian realm

D2 - Traffic Safety

ScoreDescriptionObservable Evidence
1Very PoorSpeed limits >50km/h in residential; no traffic calming; high conflict points
2PoorSpeed limits 40-50km/h; minimal calming; some conflict points unmanaged
3AdequateSpeed limits 30-40km/h; basic traffic calming; main conflict points addressed
4GoodSpeed limits 30km/h; comprehensive calming (raised tables, chicanes, narrowings); safe junction design
5Excellent20km/h zones or car-free areas; shared space design; Vision Zero principles fully applied; near-zero conflict

D3 - Perceived Safety (CPTED)

ScoreDescriptionObservable Evidence
1Very PoorPoor natural surveillance; dark areas; entrapment spots; no territorial definition
2PoorLimited surveillance from buildings; inconsistent lighting; some hidden areas
3AdequateModerate surveillance; adequate lighting (>10 lux); few hidden areas; basic territorial markers
4GoodGood natural surveillance from active uses; consistent lighting (>15 lux); clear sight lines; defined territories
5ExcellentExcellent natural surveillance; bright consistent lighting (>20 lux); no entrapment spots; clear ownership; maintained environment

D4 - Noise Environment

ScoreDescriptionObservable Evidence
1Very Poor>70 dB Lden in living areas; no noise mitigation; adjacent to motorway or rail without buffer
2Poor65-70 dB Lden; minimal mitigation; significant traffic noise throughout
3Adequate55-65 dB Lden; some noise mitigation; quiet side facades available
4Good50-55 dB Lden in most areas; effective noise mitigation; quiet courtyards and parks
5Excellent<50 dB Lden in living areas; designated quiet areas; acoustic design excellence; traffic noise eliminated from public spaces

D5 - Microclimate Comfort

ScoreDescriptionObservable Evidence
1Very PoorNo wind protection; no shade in summer; no sun access in winter; uncomfortable >50% of year
2PoorMinimal wind/shade strategy; some discomfort issues; large exposed areas
3AdequateBasic wind and shade consideration; mostly comfortable in moderate conditions
4GoodEffective wind protection; good shade provision; winter sun access; rain shelter on key routes
5ExcellentComprehensive microclimate design; wind comfort verified by CFD; summer shade >60%; winter sun optimized; thermal comfort >80% of daylight hours
Category E: IDENTITY AND CHARACTER

E1 - Legibility

ScoreDescriptionObservable Evidence
1Very PoorDisorienting; no landmarks; repetitive layout; impossible to form mental map
2PoorWeak structure; few landmarks; confusing wayfinding; limited spatial hierarchy
3AdequateBasic structure readable; some landmarks; functional wayfinding; identifiable center
4GoodClear paths, nodes, landmarks; intuitive wayfinding; distinct character areas; strong spatial hierarchy
5ExcellentHighly legible; memorable landmarks; effortless navigation; rich mental map; clear district identity within city context

E2 - Visual Appropriateness

ScoreDescriptionObservable Evidence
1Very PoorCompletely out of context; alien scale and materials; no relationship to setting
2PoorWeak contextual response; jarring scale shifts; unrelated material palette
3AdequateAcceptable contextual relationship; reasonable scale; some material connections
4GoodThoughtful contextual response; harmonious scale; complementary materials; contemporary-contextual balance
5ExcellentMasterful contextual integration; enriches the setting; confident contemporary identity rooted in place; material excellence

E3 - Human Scale

ScoreDescriptionObservable Evidence
1Very PoorOverwhelming scale; no ground-level articulation; H:W >1:5 or undefined; inhuman proportions
2PoorDominant large scale; minimal ground articulation; H:W issues; limited pedestrian-level detail
3AdequateModerate scale; some ground-level detail; H:W roughly 1:2 to 1:3; acceptable proportions
4GoodComfortable scale; good ground-level articulation; H:W 1:1 to 1:2; vertical rhythm; canopies and projections
5ExcellentExquisite human scale; rich ground-level detail and texture; H:W around 1:1; fine vertical rhythm; sense of enclosure and intimacy

E4 - Richness of Experience

ScoreDescriptionObservable Evidence
1Very PoorMonotonous; single material; no sensory variation; sterile environment
2PoorLimited variety; 1-2 materials; minimal sensory interest; repetitive
3AdequateSome variety in materials and spaces; moderate sensory interest; functional planting
4GoodRich material palette; varied spatial sequence; seasonal planting; water features or art; multi-sensory
5ExcellentExtraordinary richness; diverse materials, textures, sounds, scents; surprising spatial sequences; public art; seasonal transformation; delight

E5 - Heritage Sensitivity

ScoreDescriptionObservable Evidence
1Very PoorHeritage assets destroyed or ignored; no acknowledgment of site history
2PoorMinimal heritage response; important features lost; token gestures
3AdequateHeritage assets retained; basic setting respected; some interpretation
4GoodHeritage assets enhanced; setting improved; adaptive reuse; meaningful interpretation
5ExcellentHeritage assets celebrated; exemplary adaptive reuse; rich storytelling; archaeological sensitivity; memory of place woven into design
Category F: SUSTAINABILITY AND RESILIENCE

F1 - Environmental Performance

ScoreDescriptionObservable Evidence
1Very PoorNo energy strategy; conventional construction; no renewables
2PoorBasic code compliance only; minimal sustainability measures
3AdequateEnergy-efficient design; some renewables; meets current standards
4GoodNear-zero carbon operation; significant renewables; low embodied carbon strategy; exceeds standards
5ExcellentNet-zero or net-positive energy; comprehensive lifecycle carbon strategy; on-site renewables; Passivhaus or equivalent

F2 - Water Management

ScoreDescriptionObservable Evidence
1Very PoorConventional drainage; full surface runoff to sewer; no water strategy
2PoorBasic SUDS; limited permeable surfaces; no rainwater harvesting
3AdequateReasonable SUDS integration; some permeable surfaces (20-40%); basic rainwater collection
4GoodComprehensive SUDS; >40% permeable surfaces; rainwater harvesting; greywater recycling; bioswales
5ExcellentWater-sensitive urban design; >60% permeable; closed-loop water; sponge city principles; visible water celebration; flood-positive

F3 - Biodiversity

ScoreDescriptionObservable Evidence
1Very PoorNet biodiversity loss; no habitat provision; all hard landscape
2PoorMinimal biodiversity; ornamental planting only; no ecological strategy
3AdequateBiodiversity maintained; some native planting; basic habitat provision
4GoodBiodiversity net gain; significant native planting; green roofs; habitat corridors; ecological connectivity
5ExcellentSignificant biodiversity net gain (>20%); rich habitat mosaic; ecological corridors connected to wider network; community growing; urban forest strategy

F4 - Climate Adaptation

ScoreDescriptionObservable Evidence
1Very PoorNo climate adaptation measures; vulnerable to flooding, heat, drought
2PoorMinimal adaptation; addresses one climate risk only; reactive approach
3AdequateAddresses main climate risks; moderate heat island mitigation; basic flood resilience
4GoodComprehensive adaptation; urban heat island reduction; flood resilience; drought preparedness; cool materials
5ExcellentClimate-positive design; extensive urban forest; cool surfaces throughout; flood resilience for 1-in-100+; adaptive capacity for 2050+ scenarios

F5 - Resource Efficiency

ScoreDescriptionObservable Evidence
1Very PoorNo resource strategy; demolish-and-rebuild; no waste targets
2PoorBasic waste management; conventional materials; no lifecycle thinking
3AdequateRecycled content; construction waste targets (<15 kg/m2); some adaptable buildings
4GoodCircular economy principles; design for disassembly; significant recycled/local materials; adaptable buildings
5ExcellentFull circular economy; cradle-to-cradle materials; design for disassembly throughout; material passport; adaptable infrastructure; zero waste target

4. Quantitative Benchmarks

In addition to the qualitative scoring, check the following quantitative metrics against established benchmarks. Report each as pass/fail with the actual value.

MetricTarget RangeSource StandardNotes
Block Perimeter300-500mGehl / LEED-ND<300m may be too fine for vehicles; >500m reduces walkability
Intersection Density>100 per km2LEED-NDMeasured as 3+ leg intersections per square kilometer
Floor Area Ratio (FAR)Context-dependent: 0.5-1.5 suburban, 1.5-4.0 urban, 4.0-10.0 centralLocal zoning + Density AtlasMust relate to transit capacity and infrastructure
Green Space per Capita>9 m2/personWHO recommendationAccessible within 300m; higher targets in family neighborhoods
Parking Ratio<0.5 spaces/unit (urban), <1.0 (suburban)TOD StandardLower near high-frequency transit; includes shared parking
Height-to-Width Ratio (streets)1:1 to 1:3Gehl / AlexanderProportional enclosure; >1:4 feels exposed; <1:0.5 feels oppressive
Active Frontage>80% on primary streets, >50% on secondaryGehlMeasured as linear meters of active use / total frontage
Tree Canopy Coverage>25% of public realm areaUrban Forest StandardsAt maturity; species diversity required
Dwelling Density50-200 DU/ha (urban)Density Atlas / LEED-NDContext-dependent; must support desired service threshold
Walk + Cycle Mode Share Target>50% of tripsSustainable TransportMeasured by trip generation model or comparable precedent
Footpath Width>2.0m clear (minimum), >3.0m (preferred)Accessibility StandardsClear width excluding furniture and obstructions
Cycling Network Density>3 km/km2Dutch CROW StandardProtected or separated facilities
Public Space within 300m100% of dwellingsWHO / LEED-NDAny public space >0.1 ha within 300m walking distance
Energy PerformanceNet-zero operational carbon targetParis Agreement / LEED-NDPathway to net-zero by 2050 at minimum
Affordable Housing>20% of total unitsLocal policy dependentMix of affordable, social, and intermediate tenures

5. Output Format

Present evaluation results using the following standardized report format. Always complete every section.

# Design Evaluation Report: [Project Name]
**Location:** [City, Country]
**Scale:** [Site area in hectares] | [Approximate number of dwellings/GFA]
**Evaluator:** Claude (Urban Design Skills)
**Date:** [Current date]
**Framework Applied:** [Composite Scorecard / Gehl Deep Dive / CPTED / etc.]

---

## Overall Score: [X] / 150 - [Rating]

| Rating Band | Score Range | Description |
|-------------|------------|-------------|
| Excellent   | 120-150    | Exemplary urban design; benchmark quality |
| Good        | 90-119     | Strong design with minor improvements possible |
| Adequate    | 60-89      | Acceptable design with significant improvement opportunities |
| Poor        | <60        | Fundamental design issues requiring major revision |

---

## Category Summary

| Category | Score (/25) | Average (/5) | Rating |
|----------|------------|--------------|--------|
| A. Connectivity & Access | [X] | [X.X] | [Rating] |
| B. Vitality & Mix | [X] | [X.X] | [Rating] |
| C. Public Realm | [X] | [X.X] | [Rating] |
| D. Comfort & Safety | [X] | [X.X] | [Rating] |
| E. Identity & Character | [X] | [X.X] | [Rating] |
| F. Sustainability & Resilience | [X] | [X.X] | [Rating] |
| **TOTAL** | **[X]/150** | **[X.X]** | **[Rating]** |

---

## Detailed Scorecard

| # | Criterion | Category | Score (1-5) | Evidence / Justification | Recommendation |
|---|-----------|----------|-------------|--------------------------|----------------|
| A1 | Permeability | Connectivity | [X] | [Specific observed evidence] | [Specific improvement action] |
| A2 | Street Connectivity | Connectivity | [X] | ... | ... |
| ... | ... | ... | ... | ... | ... |
| F5 | Resource Efficiency | Sustainability | [X] | ... | ... |

---

## Quantitative Metrics Check

| Metric | Target | Actual / Estimated | Pass/Fail | Notes |
|--------|--------|--------------------|-----------|-------|
| Block Perimeter | 300-500m | [Xm] | [P/F] | [Notes] |
| Intersection Density | >100/km2 | [X/km2] | [P/F] | ... |
| ... | ... | ... | ... | ... |

---

## Radar Chart Data (for visualization)

Category A: [X.X]
Category B: [X.X]
Category C: [X.X]
Category D: [X.X]
Category E: [X.X]
Category F: [X.X]

---

## Top 5 Strengths

1. [Strength with evidence and score reference]
2. ...
3. ...
4. ...
5. ...

## Top 5 Areas for Improvement

1. [Weakness with evidence, score reference, and impact]
2. ...
3. ...
4. ...
5. ...

---

## Priority Recommendations

### Quick Wins (low cost, high impact, no redesign required)
1. [Action] - Expected improvement: [criterion] from [X] to [Y]
2. ...
3. ...

### Medium-Term Improvements (moderate cost/effort)
1. [Action] - Expected improvement: [criterion] from [X] to [Y]
2. ...
3. ...

### Structural Changes (significant redesign or investment required)
1. [Action] - Expected improvement: [criterion] from [X] to [Y]
2. ...
3. ...

---

## Benchmark Comparison

| Metric | This Project | [Exemplar 1] | [Exemplar 2] | [Exemplar 3] |
|--------|-------------|--------------|--------------|--------------|
| FAR | ... | ... | ... | ... |
| DU/ha | ... | ... | ... | ... |
| Green Space m2/pp | ... | ... | ... | ... |
| Active Frontage % | ... | ... | ... | ... |

(Select 2-3 comparable exemplars from references/benchmarks.md)

Rating Scale Interpretation:

  • Excellent (120-150): The design demonstrates best practice across most categories. Suitable as a benchmark project. Minor refinements only.
  • Good (90-119): The design is strong overall with clear strengths. Targeted improvements in weak categories would elevate it to exemplary status.
  • Adequate (60-89): The design meets basic standards but has significant room for improvement. Several categories need focused attention.
  • Poor (<60): The design has fundamental issues across multiple categories. Major redesign recommended before proceeding.

For deeper framework-specific evaluation criteria, refer to:

  • references/evaluation-criteria.md - Full criteria from Gehl, Bentley, PPS, Healthy Streets, CPTED, Universal Design, and Biophilic Design frameworks
  • references/benchmarks.md - Exemplar project metrics for benchmarking comparisons

External references:

  • Jan Gehl, "Cities for People" (2010) - 12 Quality Criteria
  • Ian Bentley et al., "Responsive Environments" (1985) - 7 Qualities
  • Project for Public Spaces, "What Makes a Successful Place?" - 4 Qualities Framework
  • Transport for London, "Healthy Streets Indicators" (2017) - 10 Indicators
  • LEED for Neighborhood Development v4.1 - Rating System
  • BREEAM Communities Technical Manual (2023)
  • CPTED Guidelines - ICA (International CPTED Association)
  • Universal Design Principles - Centre for Universal Design, NC State University
  • Terrapin Bright Green, "14 Patterns of Biophilic Design" (2014)
  • WHO Urban Green Space Recommendations
  • ITDP TOD Standard v3.1

© Abhinavbwj, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files (references) in skills/design-evaluation of Abhinavbwj/Urban-Design-Skills-Claude.

  • SKILL.md
  • references/benchmarks.md
  • references/evaluation-criteria.md

Open the folder on GitHubat commit 666327b

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in Abhinavbwj/Urban-Design-Skills-Claude, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Design Evaluation next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Design Evaluation compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Design Evaluation this skillAbhinavbwj/Urban-Design-Skills-Claude1321 repos~9.2kAutomated safety check: PassMIT
Cleanupsimstudioai/sim30k—~1.4kAutomated safety check: PassApache-2.0
122 Java Type Designjabrena/plinth446—~977Automated safety check: PassApache-2.0
Vercel Composition Patternssupabase/supabase111k59 repos~726Automated safety check: PassMIT
Swiftui View RefactorDimillian/Skills4k5 repos~2kAutomated safety check: PassMIT
RTK Rust Design Patternsrtk-ai/rtk83k—~1.9kAutomated safety check: PassApache-2.0

Similar skills

  • Cleanup

    simstudioai/sim

    Run all code quality skills — effects, memo, callbacks, state, React Query, emcn design review, url-state, comments, and test-audit — analyzing in parallel, then applying fixes sequentially

    30k GitHub stars~1.4k tokensUpdated today
    DevelopmentAuto-check passed
  • 122 Java Type Design

    jabrena/plinth

    A skill your agent uses when you need to review, improve, or refactor Java code for type design quality — including establishing clear type hierarchies, applying consistent naming conventions…

    446 GitHub stars~977 tokensUpdated today
    DevelopmentAuto-check passed
  • Official

    React composition patterns that scale. An agent skill from supabase/supabase.

    111k GitHub starsUsed in 59 repos~726 tokens
    DevelopmentAuto-check passed
  • Swiftui View Refactor

    Dimillian/Skills

    Refactor and review SwiftUI view files with strong defaults for small dedicated subviews, MV-over-MVVM data flow, stable view trees, explicit dependency injection, and correct Observation usage.

    4k GitHub starsUsed in 5 repos~2k tokens
    DevelopmentAuto-check passed
  • Describes seven Rust design patterns for the RTK CLI filter modules, with when to use each, RTK examples, and notes on when a pattern is overkill.

    83k GitHub stars~1.9k tokensUpdated today
    DevelopmentAuto-check passed
  • Effect Client Wrapper

    UsefulSoftwareCo/executor

    Pattern for wrapping third-party SDK clients (Stripe, Resend, AWS, etc.) with Effect.

    4.1k GitHub starsUsed in 1 repo~1.4k tokens
    DevelopmentAuto-check passed

More from Abhinavbwj/Urban-Design-Skills-Claude

All 18 skills in this repo
  • Cost Estimation

    Abhinavbwj/Urban-Design-Skills-Claude

    Estimate construction costs, infrastructure costs, soft costs, and total development costs for urban design projects.

    132 GitHub starsUsed in 1 repo~5k tokens
    Auto-check passed
  • Urban Calculator

    Abhinavbwj/Urban-Design-Skills-Claude

    Python computational tools for urban design metric calculations including density, FAR, walkability scoring, parking requirements, green space analysis, and block optimization.

    132 GitHub starsUsed in 1 repo~1.9k tokens
    Auto-check passed
  • Block And Density

    Abhinavbwj/Urban-Design-Skills-Claude

    Design urban blocks and optimize density using typological analysis, FAR calculations, and building configuration strategies.

    132 GitHub starsUsed in 1 repo~7.5k tokens
    Auto-check passed
  • Mobility And Transport

    Abhinavbwj/Urban-Design-Skills-Claude

    Comprehensive mobility and transport planning for urban design including trip generation, mode split targets, street network connectivity, transit planning, cycling network design, pedestrian…

    132 GitHub starsUsed in 1 repo~6.1k tokens
    Auto-check passed
  • Precedent Study

    Abhinavbwj/Urban-Design-Skills-Claude

    Research and analyze urban design precedents systematically.

    132 GitHub starsUsed in 1 repo~5.1k tokens
    Auto-check passed
  • Site Analysis

    Abhinavbwj/Urban-Design-Skills-Claude

    Conduct comprehensive multi-scale site analysis from regional to neighborhood to site scale.

    132 GitHub starsUsed in 1 repo~8.5k tokens
    Auto-check passed

Questions about Design Evaluation

What does Design Evaluation do?

Evaluate urban designs against comprehensive criteria drawn from all major global standards, certification systems, and theoretical frameworks. Design Evaluation is an agent skill from Abhinavbwj/Urban-Design-Skills-Claude. Evaluate urban designs against comprehensive criteria drawn from all major global standards, certification systems, and theoretical frameworks.

When should I use Design Evaluation?

Design Evaluation fits situations like: the user asks to evaluate a design; review a masterplan; score a proposal; critique an urban scheme.

How do I install Design Evaluation in Claude Code?

Run `npx skills add Abhinavbwj/Urban-Design-Skills-Claude --skill design-evaluation -a claude-code`. Or copy the skill folder (skills/design-evaluation in Abhinavbwj/Urban-Design-Skills-Claude) into .claude/skills/design-evaluation in your project. Claude Code loads it when a task matches its description.

How do I install Design Evaluation in Codex?

Run `npx skills add Abhinavbwj/Urban-Design-Skills-Claude --skill design-evaluation -a codex`. Or copy the skill folder (skills/design-evaluation in Abhinavbwj/Urban-Design-Skills-Claude) into .agents/skills/design-evaluation in your project. Codex loads it when a task matches its description.

Can I use Design Evaluation in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Abhinavbwj/Urban-Design-Skills-Claude --skill design-evaluation -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/design-evaluation, .gemini/skills/design-evaluation, .github/skills/design-evaluation and .opencode/skills/design-evaluation in your project.

What does Design Evaluation need to run?

SKILL.md names no scripts, command-line tools or credentials: Design Evaluation is instructions for the agent only.

Does Design Evaluation access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Design Evaluation safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Design Evaluation use?

Design Evaluation is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Design Evaluation use?

About 9.2k tokens (SKILL.md is roughly 37k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 19k tokens, read only when the agent opens those files.

What are the alternatives to Design Evaluation?

Skills that share tags, products or a category with Design Evaluation: Cleanup (simstudioai/sim, 30k stars), 122 Java Type Design (jabrena/plinth, 446 stars), Vercel Composition Patterns (supabase/supabase, 111k stars) and Swiftui View Refactor (Dimillian/Skills, 4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Design Evaluation?

Abhinavbwj (a GitHub user) maintains it in Abhinavbwj/Urban-Design-Skills-Claude, which has 132 GitHub stars. The repository holds 18 skills in this directory. The repository was last updated on March 12, 2026.

Source: Abhinavbwj/Urban-Design-Skills-Claude on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.